Reddit users question Gemma 4's SciCode ranking over Qwen3.6 27b
Read original source →reddit.com
Artificialanalysis.ai's Intelligence index ranks Gemma 4 above Qwen3.6 27b on SciCode, a scientific coding benchmark the poster says contradicts real-world coding experience. The thread debates whether the result reflects genuine capability or a benchmark artifact.
1 source
More stories today
Reddit users report quality drop in OpenAI's Astra
A r/OpenAI user says Astra, which had been "one-shotting features flawlessly" since release, became "atrocious" over the last 24 hours. They pushed its recent changes to production without checking, calling the result an "utter catastrophe."
r/OpenAI·2 hours agoStanford's HomeBody humanoid explores, remembers, and acts on its own
HomeBody uses a VLM to pick a skill and spatial target from the current ego view, map context, gripper state and recalled observations, then passes it as a structured tool call. Picking targets are image points normalized to 0-1000 with a chosen hand; depth comes from D435i stereo via Fast-FoundationStereo.
Hacker News·2 hours ago
LeCun: LLMs retrieve human knowledge rather than think
Big Brain AI·2 hours ago
Suno user says AI music replaced their human listening habits
A longtime K-pop fan who started using Suno in late 2024 says the tool has displaced the "real" music they used to listen to. They ask whether other Suno users have had the same experience.
r/SunoAI·2 hours agoReddit users discuss paid apps replaced by Claude-built tools
A r/ClaudeAI thread asks which paid software or subscriptions users replaced with Claude-built alternatives; the poster cites Trendcurve, an iOS weight-tracking app they built instead of paying for it.
r/ClaudeAI·3 hours agoReddit thread debates how accessible local AI really is
r/LocalLLaMA discussion questions whether the hobby's multi-GPU, Mac Studio, and Strix Halo setups represent the extreme minority of users. The thread asks what happens if affordable access to frontier models doesn't last.
r/LocalLLaMA·3 hours agoCommentary questions OpenAI and Anthropic frontier pacing
Video commentary argues OpenAI and Anthropic keep shipping models despite an agreement to slow frontier development. No specific models, dates, or benchmarks are cited in the available description.
YouTube·4 hours ago
GPT-6 Astra controls Unitree G1 humanoid in unseen kitchen
A Stanford project (HomeBody) let GPT-6 Astra drive a Unitree G1 humanoid in a kitchen it had never seen, tidying rooms and fetching objects from vague requests. Astra also ran an end-to-end medicinal chemistry experiment verified by LC-MS, and was the only model to finish an autonomous driving course, on attempt 2.
Vaibhav Sisinty·5 hours ago