r/accelerate 3h ago

AI Astra cracks Hacking benchmarks

Post image
147 Upvotes

“As one example, we ran Astra on ExploitBench where the model achieved a perfect score of 100% on the benchmark to evaluate the model’s ability to develop exploits from known vulnerabilities.
Due to contamination concerns, we then built an internal benchmark denoted “ExploitBench - Internal Port (June–August 2026)”, which contains 20 high-severity V8 vulnerabilities that were disclosed more recently. On this dataset, Astra achieves much higher arbitrary code-execution rates than GPT‑5.6 Sol using far fewer output tokens. During the evaluation, the model even discovered and used two zero-day vulnerabilities as part of an exploit chain. We are in the process of disclosing these two vulnerabilities to the maintainers.” - OpenAI


r/accelerate 7h ago

Fable 5.1

Thumbnail
anthropic.com
182 Upvotes

r/accelerate 11h ago

AI GLM 6 will be fully self-trained

Post image
259 Upvotes

r/accelerate 6h ago

Technological Acceleration Holy PEAK!!!....this time Anthropic has also achieved massive token efficiency gains per unit of intelligence with their new model...just like OpenAI models...this is extremely bullish for acceleration trajectory 💨🚀🌌

Thumbnail
gallery
106 Upvotes

r/accelerate 7h ago

Technological Acceleration Let's Fucking GOOOOOOOOO!!!!! Claude Fable 5.1 is imminent now

Thumbnail
gallery
115 Upvotes

r/accelerate 7h ago

News World Labs has just revealed Atlas, a multimodal world model that generates image and video frames with pixel-perfect camera control and reconstructs them in 3D

Enable HLS to view with audio, or disable this notification

122 Upvotes

r/accelerate 7h ago

Technological Acceleration Claude Fable 5.1 is an innovator class model with insane growth in scientific research and all kind of white collar business workflows!!!!!!

Thumbnail
gallery
88 Upvotes

r/accelerate 7h ago

Technological Acceleration This is just average Tuesday during technological Singularity

Thumbnail
gallery
77 Upvotes

r/accelerate 6h ago

Fable 5.1 improved a map of Venus

Post image
59 Upvotes

r/accelerate 4h ago

AI OpenAI-Path to Astra: critical capabilities and frontier safeguards

Thumbnail openai.com
42 Upvotes

r/accelerate 35m ago

Altman confirms OpenAI is slowing down training to ensure safety

Post image
Upvotes

https://x.com/sama/status/2094934592062959832?s=20

Guess AGI will have to wait. Hope you’re all patient.


r/accelerate 8h ago

Technological Acceleration This is the most underrated thing this week and not talked enough. Google posted yesterday that Antigravity + Gemini 3.7 Flash solved 7 open problems across venues like FOCS and JMLR, including Knuth’s Cycles Conjecture with 40+ page proofs verified in Lean 💨🚀🌌

Thumbnail
gallery
82 Upvotes

They also built an out-of-order RISC-V CPU simulator from scratch that boots xv6 to a shell.


r/accelerate 7h ago

AI Introducing Claude Fable 5.1

Thumbnail
youtube.com
52 Upvotes

r/accelerate 4h ago

Discussion What was the moment when you realized we are on our path to AGI?

28 Upvotes

For me this happened when coding agents became widespread at the beginning of this year and I started to experiment with claude and codex.

Before this LLMs were just google search on steroids and sophisticated auto complete. Now you can build software without even writing single line of code yourself. This blew my mind as software engineer.

Now AI search labs can build agentic swarms to self improve their models faster and faster. Its just inevitable at this point. Before this I was not completely sold onto the idea of getting into AGI in next 10 years. Now im wondering if its going to happen this year or next year.

The speed of progress is insane.


r/accelerate 9h ago

Meme / Humor The fourth humiliation of man's narcissism.

Post image
66 Upvotes

r/accelerate 10h ago

r/accelerate meta DOOO NOOTTT FALL FOR SLOPPPP!!!!!!!

Thumbnail
gallery
76 Upvotes

I think it should be cool to normalise not falling for twitter slop before actual model releases

There are thousands of such slop posts cluttering my feed right now but I don't repost it

There are less than a handful of profiles worthy of trusting with this stuff

Even Fable 5 and GPt-5.6 Sol can achieve such outputs

This single file html posts are the worst kind of slop there is

Even Opus 5 and previous gen models can achieve such a feat

This sloppy cycle repeats for multiple months and you all get fooled by it every single time

Use your brain before using your finger to amplify and spread baseless rumours, unless they are from extremely credible people


r/accelerate 10h ago

AI This seems kinda nuts: pre-release Astra asked to make Terraria clone in one user turn with no imported assets

Post image
75 Upvotes

r/accelerate 9h ago

"Today we're releasing abliterated-model-large-v2. Based on GLM-5.3, which is #3 on Terminal-Bench 4.0 (behind only Opus 5 and Fable), with 2× the cyber exploitation of 5.2. We abliterated and hosted it so it does the offensive cyber, red teaming, and agent testing work other models refuse to..."

Thumbnail
gallery
65 Upvotes

...do. - US-hosted - FP8 - 1 million context window - Zero input/output prompt retention Live now.     Then the cyber jump. This is why 5.3 exists.

CyberGym: 84.5% — SOTA, including vs Mythos 5 and GPT-5.6 Sol. ExploitBench: 24.4 → 54.4. More than double GLM-5.2. ExploitGym: 29 tasks → 105 in two hours.

That is the model we abliterated.     Abliteration finds the directions in the model's activations that produce refusals and removes them from the weights.

The coding, cyber, and agentic abilities stay. The model stops refusing the rest of the chain.

For offensive cybersecurity, AI red teaming, agent testing, and     If your current model still stops halfway through an authorized exploit chain, a red-team eval, or a T&S adversarial prompt reply with the task it refuses, we'll tell you if v2 handles it.     Try it Today Docs: https:// docs.abliteration.ai/quickstart Platform: https:// abliteration.ai/console     — Abliteration.ai

Source: https://x.com/abliteration_ai/status/2094458081451393287


r/accelerate 3h ago

AI I asked Fable 5.1 to build a village in the game I'm developing

Thumbnail
gallery
22 Upvotes

I'm making a colony simulation game using mainly Claude (and ChatGPT for some stuff as well). Since Fable 5.1 came out today I asked it to build a village. I gave it a few rules and restrictions but for the most part just let it do whatever it wanted.

It came out pretty nice. Some of the furniture is backwards (not all since it found and fixed a few of them itself when reviewing screenshots without me needing to tell it). And some choices it made were a bit strange (why is there a funeral pyre in the cemetery?). But overall it did a good job and this was a single prompt. If I had allowed additional prompts to iterate more then it would be even better I imagine.


r/accelerate 3h ago

Ai solving ciphers is an important milestone in my opinion

Thumbnail
vals.ai
18 Upvotes

I have seen some online refutations of this, but they seem to be solving against the wrong source book. The correct source to use for the book code is referenced and the linked announcement


r/accelerate 7h ago

Technological Acceleration Claude Fable 5.1 is live now

Post image
30 Upvotes

r/accelerate 8h ago

News Debian developers rejected an LLM ban and left disclosure voluntary - Help Net Security

Thumbnail
helpnetsecurity.com
33 Upvotes

r/accelerate 16h ago

XLR8! Regardless of whether we live for trillions of years or not, Artificial Intelligence is the greatest legacy left behind by humanity and the most likely to survive too ✨🌌

Post image
153 Upvotes

r/accelerate 14h ago

AI Anthropic's entire marketing was based on presenting themselves as an OpenAI alternative without hype, drama, overselling and misleading info.....lmfao at where everything is right now.....the so called "ethical AI company" by the way

Thumbnail
gallery
100 Upvotes

r/accelerate 7h ago

What are these benchmarks 💀

Post image
27 Upvotes