Image 1 — "This is interesting: Claude is already achieving roughly twice the protein-design hit rate of conventional human-led workflows. 27% hit rate in autonomous protein binder design, roughly twice the typical 10–15% rate reported in the field. Working from one expert-written protocol, Claude..."
Image 2 — "This is interesting: Claude is already achieving roughly twice the protein-design hit rate of conventional human-led workflows. 27% hit rate in autonomous protein binder design, roughly twice the typical 10–15% rate reported in the field. Working from one expert-written protocol, Claude..."
Image 3 — "This is interesting: Claude is already achieving roughly twice the protein-design hit rate of conventional human-led workflows. 27% hit rate in autonomous protein binder design, roughly twice the typical 10–15% rate reported in the field. Working from one expert-written protocol, Claude..."
Image 4 — "This is interesting: Claude is already achieving roughly twice the protein-design hit rate of conventional human-led workflows. 27% hit rate in autonomous protein binder design, roughly twice the typical 10–15% rate reported in the field. Working from one expert-written protocol, Claude..."
Image 5 — "This is interesting: Claude is already achieving roughly twice the protein-design hit rate of conventional human-led workflows. 27% hit rate in autonomous protein binder design, roughly twice the typical 10–15% rate reported in the field. Working from one expert-written protocol, Claude..."
▲ 8 r/ProAI

"This is interesting: Claude is already achieving roughly twice the protein-design hit rate of conventional human-led workflows. 27% hit rate in autonomous protein binder design, roughly twice the typical 10–15% rate reported in the field. Working from one expert-written protocol, Claude..."

> Many drugs work by binding to a specific target in the body and blocking or changing what it does. An important first step in the drug development process is designing a molecule that can bind tightly to its target. Traditionally, that's meant weeks or months of expert work per https://t.co/CGCNTNaKBq >   > — Anthropic

Source: https://x.com/AnthropicAI/status/2089842387845804246


> This is interesting: Claude is already achieving roughly twice the protein-design hit rate of conventional human-led workflows. 27% hit rate in autonomous protein binder design, roughly twice the typical 10–15% rate reported in the field. > > Working from one expert-written protocol, Claude designed binders against 14 of 15 measurable targets. Independent labs confirmed that 354 of 1,320 designs bound successfully. > > Depending on the setup, Claude’s hit rate ranged from 22.6% to 35.1%. Its top-ranked design bound in 49% of campaigns. > > This is not yet fully autonomous drug research, but it is another important building block in that direction. >   >   > — Chubby

Source: https://x.com/kimmonismus/status/2089852014331117694

u/stealthispost — 16 hours ago

"This is interesting: Claude is already achieving roughly twice the protein-design hit rate of conventional human-led workflows. 27% hit rate in autonomous protein binder design, roughly twice the typical 10–15% rate reported in the field. Working from one expert-written protocol, Claude..."

> Many drugs work by binding to a specific target in the body and blocking or changing what it does. An important first step in the drug development process is designing a molecule that can bind tightly to its target. Traditionally, that's meant weeks or months of expert work per https://t.co/CGCNTNaKBq >   > — Anthropic

Source: https://x.com/AnthropicAI/status/2089842387845804246


> This is interesting: Claude is already achieving roughly twice the protein-design hit rate of conventional human-led workflows. 27% hit rate in autonomous protein binder design, roughly twice the typical 10–15% rate reported in the field. > > Working from one expert-written protocol, Claude designed binders against 14 of 15 measurable targets. Independent labs confirmed that 354 of 1,320 designs bound successfully. > > Depending on the setup, Claude’s hit rate ranged from 22.6% to 35.1%. Its top-ranked design bound in 49% of campaigns. > > This is not yet fully autonomous drug research, but it is another important building block in that direction. >   >   > — Chubby

Source: https://x.com/kimmonismus/status/2089852014331117694

u/stealthispost — 16 hours ago

"Strawberries. Indoors. Stacked to the ceiling. Now add drones flying between the rows. In the 4D Bios setup, drones move through vertical shelves and capture high-resolution images of leaves and fruit. Instead of manual inspection, plant health becomes a data stream: > drones scan them all the..."

> ...time and turn ripeness, disease, and growth into real data that systems can act on. That’s when agriculture starts looking a lot like advanced manufacturing. Video: https:// youtu.be/cPjnooZUetI —- Weekly robotics and AI insights. Subscribe free: http:// 22astronauts.com >   >   > — Ilir Aliu

Source: https://x.com/IlirAliu_/status/2089984758486835479

u/stealthispost — 16 hours ago
▲ 4 r/ProAI

"Very interesting development. Likely a result of new requirement of log-in to access old Reddit: https:// arstechnica.com/gadgets/2026/0 6/reddit-will-require-you-to-log-in-to-use-old-reddit-com/ …"

> looks like reddit is almost wiped from chatgpt sources > > the query fanout changes had a big impact > > and the past couple of days it seems to be almost completely removed from prompt responses > > https://t.co/oCGm9M0yPO https://t.co/c2GJF2apG7 >   > — Klaas

Source: https://x.com/forgebitz/status/2089708381351059924


— Kevin Bankston

Source: https://x.com/KevinBankston/status/2089768281892638774

u/stealthispost — 16 hours ago

"Very interesting development. Likely a result of new requirement of log-in to access old Reddit: https:// arstechnica.com/gadgets/2026/0 6/reddit-will-require-you-to-log-in-to-use-old-reddit-com/ …"

> looks like reddit is almost wiped from chatgpt sources > > the query fanout changes had a big impact > > and the past couple of days it seems to be almost completely removed from prompt responses > > https://t.co/oCGm9M0yPO https://t.co/c2GJF2apG7 >   > — Klaas

Source: https://x.com/forgebitz/status/2089708381351059924


— Kevin Bankston

Source: https://x.com/KevinBankston/status/2089768281892638774

u/stealthispost — 16 hours ago

Robotic firefighters sent out to inspect a self-landing booster right after touchdown. just normal 2026 stuff

> It does feel like there's a spark there. >   > — 星落 >   >   > lol. Yes! Both robots also sprayed fire suppressant. It seems to be part of the design assembly. >   > — HG 阿聻 𓆣 𓇽

Source: https://x.com/ohgkg/status/2089880928294474065

u/stealthispost — 16 hours ago

"MISSION SUCCESS | ZhuQue-3 Y2 Reusable Launch Vehicle Achieved Full Success in Orbital Insertion and First-Stage Recovery On August 19, 2026, at 07:35 (UTC+8), the ZhuQue-3 (ZQ-3) Y2 reusable launch vehicle lifted off from the Dongfeng Commercial Space Innovation Pilot Zone. Approximately 137..."

> MISSION SUCCESS | ZhuQue-3 Y2 Reusable Launch Vehicle Achieved Full Success in Orbital Insertion and First-Stage Recovery > > On August 19, 2026, at 07:35 (UTC+8), the ZhuQue-3 (ZQ-3) Y2 reusable launch vehicle lifted off from the Dongfeng Commercial Space Innovation Pilot Zone. Approximately 137 seconds after lift-off, the first and second stages separated. The second stage continued its flight and successfully delivered the Honghu 03 satellite, independently developed by Hongqing Technology, into its designated orbit. At approximately 07:41, the first stage performed a successful soft touchdown at the LandSpace Landing Site#1 in Minqin County, Gansu Province, following the planned trajectory — marking the full success of the flight test mission. >   >   > This mission marks China's first-ever successful recovery attempt of the first stage of an orbital-class launch vehicle using landing legs, and China's first successful booster recovery on land. It is a pivotal flight test for ZhuQue-3 as it transitions from the technological >   >   > After stage separation, the first stage completed a series of critical maneuvers during a planned coast phase: high-altitude attitude reorientation, powered deceleration via re-entry burn, aerodynamic gliding, landing burn, and landing leg deployment with touchdown attenuation — >   >   > Building on flight data and engineering experience from the ZQ-3 Y1 flight test, ZQ-3 Y2 incorporates multiple targeted optimizations across key phases of recovery flight. The landing propulsion scheme reduces the number of engines used in the landing burn, further simplifying >   >   > Leveraging the characteristics of this mission, ZQ-3 Y2 also carried LandSpace's independently developed pyrotechnics-free stacking Hold-Down and Release Mechanism (HDRM), further validating its performance in real mission conditions and providing strong support for future >   >   > ZhuQue-3 is LandSpace's independently developed new-generation reusable LOX–methane launch vehicle, with reusability built into its overall design, propulsion system, return control, stainless steel vehicle structure manufacturing, and ground support. This mission lays a solid >   >   > — LandSpace

Source: https://x.com/LandSpace_Tech/status/2089877715331809546

u/stealthispost — 16 hours ago
▲ 2 r/ProAI

"OpenAI's president just went on CNBC and read our April thesis back to us Brockman: "Compute is really becoming the new oil, the new limited resource of the AI age" We ran it April 4. "Oil is scarce because of war. Tokens are scarce because of physics" https://..."

> ...bepresearch.substack.com/p/the-token-do llar … He even brought the proof for the next one. $2,000 of compute, 10 open math problems. Proofs verify for free. Biology needs a bench. The moat is the measurement https:// bepresearch.substack.com/p/the-next-inf lection-is-the-lab … New oil, old receipt >   > — Ben Pouladian >   >   > he's copying you Ben >   > — Alex A.C. >   >   > All good > @gdb > we can talk anytime. Full speed >   > — Ben Pouladian

Source: https://x.com/benitoz/status/2089392813758972149

u/stealthispost — 17 hours ago
▲ 239 r/ProAI+1 crossposts

"Qwen3.8 27B wrote a playable first person shooter start to finish on 2 used 3090s in my apartment. No cloud. No API key. No subscription. 687 steps. 87.9M tokens in, 822k out. 5 hours 11 minutes of model time plus 58 minutes of tool calls. The agent loop never broke once. And it plays. Enemies..."

> ...spawn and push you, the gun kicks, shadows stretch across the whole block while the sun goes down behind the towers. I sat there clearing waves instead of grading the output. Same shooter prompt I threw at the frontier models a few weeks ago. That time the tokens went to somebody elses datacenter. This time nothing left the flat. 87.9 million tokens through my own cards. On an API that run has a price tag. Here it has an electricity bill. 60 tok/s all the way through. Slower than frontier, and it stops mattering when the thing works through the night while you sleep. Local models were a toy 18 months ago. This one finished a game. >   >   > Play now! Choose Build 2 Qwen3.8 27B > > > https:// > alesha-pro.github.io/bench-portal/ >   >   > — Alexey Fateev

Source: https://x.com/superalesha/status/2089126766854238421

u/stealthispost — 22 hours ago
▲ 4 r/ProAI

"A "refusal-removed" version of Qwen3.8-27B can now run locally on Apple Silicon. Even its creators warn that it can provide malware, fraud and weapons instructions on demand. It was released as an MLX build in 2, 4, 6 and 8-bit versions. The uploader claims its 4/6/8-bit tests produced zero..."

> We just shipped our official Qwen 3.8 27B Uncensored MLX build. Local. Uncensored. For🍎 > > 2-bit, 4-bit, 6-bit & 8-bit — pick your poison based on RAM and speed. > > No CUDA. No cloud. Just your Mac and the weights. Have fun! > https://t.co/b3gXsHeSdk >   > — OrcaRouter 🐳

Source: https://x.com/OrcaRouter/status/2089385980080148726


> A "refusal-removed" version of Qwen3.8-27B can now run locally on Apple Silicon. > > Even its creators warn that it can provide malware, fraud and weapons instructions on demand. > > It was released as an MLX build in 2, 4, 6 and 8-bit versions. The uploader claims its 4/6/8-bit tests produced zero refusals while preserving vision, reasoning and tool-calling across a 262K-token context. > > The Qwen 27B Model is a very capable model. This is the first time I've really seen the immediate dangers in a tangible way. > > We need a societal discussion about this. >   > — Chubby >   >   > how long did it take to get this running locally? >   > — tan >   >   > it runs locally >   > — Chubby

Source: https://x.com/kimmonismus/status/2089763435865088508

u/stealthispost — 1 day ago

"A "refusal-removed" version of Qwen3.8-27B can now run locally on Apple Silicon. Even its creators warn that it can provide malware, fraud and weapons instructions on demand. It was released as an MLX build in 2, 4, 6 and 8-bit versions. The uploader claims its 4/6/8-bit tests produced zero..."

> We just shipped our official Qwen 3.8 27B Uncensored MLX build. Local. Uncensored. For🍎 > > 2-bit, 4-bit, 6-bit & 8-bit — pick your poison based on RAM and speed. > > No CUDA. No cloud. Just your Mac and the weights. Have fun! > https://t.co/b3gXsHeSdk >   > — OrcaRouter 🐳

Source: https://x.com/OrcaRouter/status/2089385980080148726


> A "refusal-removed" version of Qwen3.8-27B can now run locally on Apple Silicon. > > Even its creators warn that it can provide malware, fraud and weapons instructions on demand. > > It was released as an MLX build in 2, 4, 6 and 8-bit versions. The uploader claims its 4/6/8-bit tests produced zero refusals while preserving vision, reasoning and tool-calling across a 262K-token context. > > The Qwen 27B Model is a very capable model. This is the first time I've really seen the immediate dangers in a tangible way. > > We need a societal discussion about this. >   > — Chubby >   >   > how long did it take to get this running locally? >   > — tan >   >   > it runs locally >   > — Chubby

Source: https://x.com/kimmonismus/status/2089763435865088508

u/stealthispost — 1 day ago
▲ 3 r/ProAI

"Scaling self-verification with DeepSeek V4 Flash beats Claude Fable 5 on Terminal-Bench 2.1, while being 11x cheaper As open-source models become more capable, they can now generate large numbers of high-quality candidate solutions and verify their own outputs at very low cost. For example, we..."

> How can we extract richer signals from AI Feedback? > > Introducing LLM-as-a-Verifier✨— a simple verification scaling framework that achieves SOTA on agentic benchmarks 🚀 > > The key idea: > - Use fine-grained scoring granularity (e.g., 1-20 instead of the standard 1-5 scale) > - Take https://t.co/0sCeAwcar1 >   > — Jacky Kwok

Source: https://x.com/jackyk02/status/2074969820739805275


> Scaling self-verification with DeepSeek V4 Flash beats Claude Fable 5 on Terminal-Bench 2.1, while being 11x cheaper > > As open-source models become more capable, they can now generate large numbers of high-quality candidate solutions and verify their own outputs at very low cost. > > For example, we find that sampling just 5 solutions with DeepSeek V4 Flash and ranking them using the same model with LLM-as-a-Verifier can lead to a significant boost in accuracy (79% → 88%), outperforming closed frontier models on Terminal-Bench. > > Try it out today: > https:// > github.com/llm-as-a-verif > ier/llm-as-a-verifier#self-verification-terminal-bench-21 > … > > More on verification scaling in my previous post. >   > — Jacky Kwok >   >   > Is there an OpenCode plugin for this to try it out with Deepseek v4 flash? >   > — Shahbaz Ahmed >   >   > We’ll be releasing a harness on top of LLM-as-a-Verifier later this month :) >   > — Jacky Kwok

Source: https://x.com/jackyk02/status/2089421448784023553

u/stealthispost — 1 day ago

"Scaling self-verification with DeepSeek V4 Flash beats Claude Fable 5 on Terminal-Bench 2.1, while being 11x cheaper As open-source models become more capable, they can now generate large numbers of high-quality candidate solutions and verify their own outputs at very low cost. For example, we..."

> How can we extract richer signals from AI Feedback? > > Introducing LLM-as-a-Verifier✨— a simple verification scaling framework that achieves SOTA on agentic benchmarks 🚀 > > The key idea: > - Use fine-grained scoring granularity (e.g., 1-20 instead of the standard 1-5 scale) > - Take https://t.co/0sCeAwcar1 >   > — Jacky Kwok

Source: https://x.com/jackyk02/status/2074969820739805275


> Scaling self-verification with DeepSeek V4 Flash beats Claude Fable 5 on Terminal-Bench 2.1, while being 11x cheaper > > As open-source models become more capable, they can now generate large numbers of high-quality candidate solutions and verify their own outputs at very low cost. > > For example, we find that sampling just 5 solutions with DeepSeek V4 Flash and ranking them using the same model with LLM-as-a-Verifier can lead to a significant boost in accuracy (79% → 88%), outperforming closed frontier models on Terminal-Bench. > > Try it out today: > https:// > github.com/llm-as-a-verif > ier/llm-as-a-verifier#self-verification-terminal-bench-21 > … > > More on verification scaling in my previous post. >   > — Jacky Kwok >   >   > Is there an OpenCode plugin for this to try it out with Deepseek v4 flash? >   > — Shahbaz Ahmed >   >   > We’ll be releasing a harness on top of LLM-as-a-Verifier later this month :) >   > — Jacky Kwok

Source: https://x.com/jackyk02/status/2089421448784023553

u/stealthispost — 1 day ago
▲ 23 r/ProAI

"i don’t know who needs to hear this but qwen 3.8 27b is ranked ABOVE: - gpt 5.3 - gemini 3.1 pro - opus 4.6 all of which were state of the art 6 MONTHS AGO AND IT RUNS ON A LAPTOP"

> have this running on my 5090 right now and im getting 200 tk a second. i feel like I'm literally playing with magic >   > — Alex Finn >   >   > you can either buy anthropic for 2 trillion dollars or a used 3090 gpu for $1,500 > > only one of those will refuse your prompts >   > — Udi Wertheimer

Source: https://x.com/udiWertheimer/status/2089421927085400203

u/stealthispost — 2 days ago

"i don’t know who needs to hear this but qwen 3.8 27b is ranked ABOVE: - gpt 5.3 - gemini 3.1 pro - opus 4.6 all of which were state of the art 6 MONTHS AGO AND IT RUNS ON A LAPTOP"

> have this running on my 5090 right now and im getting 200 tk a second. i feel like I'm literally playing with magic >   > — Alex Finn >   >   > you can either buy anthropic for 2 trillion dollars or a used 3090 gpu for $1,500 > > only one of those will refuse your prompts >   > — Udi Wertheimer

Source: https://x.com/udiWertheimer/status/2089421927085400203

u/stealthispost — 2 days ago
▲ 5 r/ProAI

"18x improvement in intelligence per joule in 16 months."

> hard agree with @amasad —@JonSaadFalcon and my research indicates that intelligence efficiency (intelligence per watt) is rapidly improving and we will definitely not need data center scale compute to run agi! > > links to research in comments below 👇 >   > — Avanika Narayan

Source: https://x.com/Avanika15/status/2089028986932470156


— Amjad Masad

Source: https://x.com/amasad/status/2089069905375351169

u/stealthispost — 3 days ago

"18x improvement in intelligence per joule in 16 months."

> hard agree with @amasad —@JonSaadFalcon and my research indicates that intelligence efficiency (intelligence per watt) is rapidly improving and we will definitely not need data center scale compute to run agi! > > links to research in comments below 👇 >   > — Avanika Narayan

Source: https://x.com/Avanika15/status/2089028986932470156


— Amjad Masad

Source: https://x.com/amasad/status/2089069905375351169

u/stealthispost — 3 days ago

"HopTo’s Hop-1 is a hybrid hopping-and-flying robot built for rough terrain. It looks weird at first, but the logic is actually pretty smart. Flying all the time burns a ton of energy just fighting gravity, so they let the robot spend most of its time bouncing on a passive spring instead. The..."

> ...rotors mainly steer, put back the energy lost on each hop, and only fully take over when the robot actually needs to fly. HopTo says hopping uses around 25% of the energy of continuous flight. So the whole idea is pretty simple: stay on the ground as much as possible, and only fly when you have to. >   >   > One thing that’s easy to miss: hopping actually makes state estimation pretty nasty. > > The robot spends most of each jump close to free fall, so the accelerometer can’t reliably use gravity as a reference for roll/pitch like a normal drone does. Then it hits the ground, gets a >   >   > — Eren Chen

Source: https://x.com/ErenChenAI/status/2089029735128957191

u/stealthispost — 3 days ago