InternLM releases Intern-S2-Preview-397B model
InternLM released the Intern-S2-Preview-397B, a 397-billion parameter model under preview on HuggingFace. The model is already trending with community interest.
Daily AI Briefing
The 64 stories that mattered in AI, curated and summarized from dozens of sources by AIBriefs.
InternLM released the Intern-S2-Preview-397B, a 397-billion parameter model under preview on HuggingFace. The model is already trending with community interest.
Paper scales reinforcement learning with verifiable rewards (zero RL) to a trillion parameters, leading to emergent reasoning capabilities. It elicits chain-of-thought reasoning without human-annotated data.
Amazon Quick targets the 60% of sales rep time lost to low-value tasks like CRM updates and prospect research. The agentic AI teammate automates email drafting, context-switching, and research to free up selling time.
Seedance 2.5 generates 30-second 4K videos from up to 50 multimodal references. A beta long-video mode extends output to 180 seconds. The model will roll out on CupCut and partner apps.
The Ring-2.6 model family, released under an MIT license, reportedly matches frontier performance on benchmarks like ARC-AGI-v2 and AIME. The trillion-parameter model is detailed in arXiv report 2606.15079.
A study tracking 4 million applications across 150 employers found 26% of Black and 15% of Asian applicants faced AI-driven racial bias per the EEOC's four-fifths rule. If minority candidates had been recommended at the same rate, 40,000 more applications would have advanced.
BrainCo demonstrated a brain-computer interface platform enabling neural control of robots, after over a decade of development. The system, shown at the 2026 World AI Conference, also includes an AI-powered bionic hand with independent finger control.
Apple ML Research proposes RayRoPE, a positional encoding for multi-view transformers that encodes patches uniquely and allows SE(3)-invariant attention. The method can adapt to scene geometry.
The chip, codenamed Frozen v2, reportedly targets 6–10× more tokens per watt than Google's newest TPUs. Deployment is planned as early as 2028 as Google seeks to address compute shortages. Alphabet shares rose on the news.
At VB Transform 2026, LangChain's Harrison Chase and leaders from Conviva and CoreWeave argued that perfect single-agent conversations can mask broken products. They advocated moving from scoring individual traces to comparing user cohorts against baselines.
The bill would require compensation for human authors when AI uses their works and ban AI imitation of their styles. If passed, Indonesia would be the first Southeast Asian country to explicitly incorporate AI into its copyright framework.
Apple ML Research introduces LVSum, a human-annotated benchmark for long video summarization that requires both semantic and temporal grounding. It challenges multimodal large language models to maintain temporal fidelity over extended durations.
Jeff Bezos invests in CuspAI, an AI startup partnered with Nvidia to discover new chipmaking materials. The company is part of a new wave of AI firms tackling physical-world challenges.
An analysis on unslop.run finds that over 30% of new arXiv submissions appear to be AI-written. The detection method identifies text likely generated by language models.
Snyk uncovered 241 vulnerabilities in a game's code that an earlier agentic security pass by Fable had missed. Steve Yegge discusses permissions, provenance, and agent supply chain risks.
Apple proposes Length Value Model (LVM) for fine-grained token-level length control in autoregressive models. Unlike coarse-grained approaches, LVM explicitly models generation length during pretraining to optimize inference cost and reasoning performance.
The cluster is the first dedicated GPU infrastructure for the YC community, allowing startups to access compute with a few weeks' commitment instead of a typical 24-month contract. The partnership aims to speed up AI development for YC portfolio companies.
Neo raised $100 million in seed and Series A funding from Andreessen Horowitz and Bessemer Venture Partners. The company aims to help enterprises control and secure their AI software deployments.
Nativ is a new tool that enables running frontier open-source models locally on macOS. It is available now and appears to support multiple models for offline inference.
In 2023, researchers found thousands of Ray clusters exposed on the public internet, with data worth over $1B at risk. Lovina Dmello from NVIDIA discusses how the LLM stack suffers from security flaws similar to databases from 2008, emphasizing that default authentication is often missing.
Over the past nine months, Amazon, Microsoft, and Google introduced enterprise agent platforms that share core components: runtime, memory, tool gateway, identity, observability, and governance. This convergence enables agent portability across the three clouds.
Latency budget for voice agents is 200 milliseconds, far tighter than chat agents' seconds. The talk covers barge-in handling and turn-taking to avoid user frustration.
Bee wearable captures ~10 million tokens per year, learning everything about a user within a week. Korshakov explains the design guarantee that no one else can access the recorded data.
Diane Lin of Datadog argues that LLM inconsistency is a critical product flaw, especially in high-stakes fields like cybersecurity. She provides strategies to mitigate flip-flopping and build trust in agent outputs.
Ray 2.55 introduces official, first-class support for Google Cloud TPUs, enabling distributed Python workloads via familiar Ray APIs. Multi-host TPU slices must be kept together, handled by Ray's placement groups.
AI-powered network technology doubles data transmission capacity for telecom operators. First major announcement since Nvidia took a stake in Finnish network equipment maker Nokia nine months ago.
N.E.K.O. is an open-source AI companion that continuously observes desktop activity and initiates conversations. Showcased at WAIC 2026 as part of the "Catgirl Plan" ecosystem.
The post covers AI transparency practices including data provenance, model explainability, and governance frameworks. It emphasizes building user trust through clear documentation and ethical data handling.
Daily AI token calls in China reached 140 trillion by March 2026, a more than 1,000-fold increase from roughly 100 billion in early 2024. The figure was cited by CAICT deputy head Wei Liang in a CCTV Finance program preview.
26% of top security executives are considering leaving due to heightened job pressures from rapid AI adoption. The Dark Reading article highlights the strain on CISOs managing AI risk.
Project Indigo can now remove any background from photos snapped in the app. An AI feature also provides constructive critique on composition and lighting.
AWS AI Blog post demonstrates integrating vision capabilities with Amazon Bedrock and MCP servers to build agentic systems. The approach bridges the gap between systems that see, think, and act, simplifying complex integrations.
In a talk at AI Engineer, Ornella Bahidika and Joel Allou show a voice tutor where the LLM does not decide lesson timing, correctness, or next steps—a harness orchestrates while the LLM just generates responses. They argue engineers should avoid letting the LLM drive multi-step agent flows.
Seeddream 5.0 Pro accepts up to 10 reference images and generates infographics, UI mockups, and ads with readable text. It is compared to GPT Image 2 for design work.
YouTube updated its monetization policies to define which AI-generated and low-quality videos are ineligible for ad revenue. The clarification targets content deemed harmful or upsetting, aiming to reduce low-effort AI slop.
The largest probabilistic computer ever built uses thermal noise to perform computations. It can solve complex optimization and sampling problems, potentially accelerating AI and machine learning workloads.
A $400 million chip-backed loan by early GPU financiers funds inference chip infrastructure. The deal signals a shift toward specialized AI inference hardware as demand grows.
Databricks outlines three ways AI can transform retail, travel, and consumer goods: improving customer experience, optimizing supply chains, and enhancing operational efficiency. The article emphasizes the importance of a unified data platform for enabling these AI applications.
NVIDIA NemoClaw, a collection of open blueprints for autonomous agents, enables context-aware video AI agents to integrate with enterprise systems like content management, messaging, and databases. The approach moves from analysis to action by orchestrating blueprints for structured reports and multistep workflows.
Researchers can receive up to $50,000 in Claude credits over six months. Two tracks: one for basic research and one for early-stage biotechs accelerating clinical development for rare diseases.
Couchbase's Capella iQ uses a multi-model AI architecture on Amazon Bedrock to generate database queries and recommend indexes. The system supports multi-turn conversational workflows by routing requests to different LLMs.
Cursor explores how agent swarms coordinate multiple AI models and the cost implications of scaling such systems. The blog post discusses the economic trade-offs and practical benefits of using model swarms in development workflows.
The system handles 15,000 questions daily from Slack, GitHub, and Confluence. Built on Cerebras's hardware, it retrieves from structured and unstructured data.
Entry-level associate hiring at AmLaw 200 firms remained flat at ~7,400 per year from 2022-2025 despite combined revenue growth of tens of billions. Lateral hiring rose, exceeding newbie hires in 2025, suggesting AI may be reducing demand for junior lawyers.
Elizabeth Stone, Netflix Chief Product and Technology Officer, shares insights on how AI is transforming product management and engineering roles. The conversation explores the evolving skill sets needed and the balance between automation and human creativity.
The director of the Trump administration's AI safety agency (CAISI) has resigned after three months. Arvind Raman, director of NIST, will serve as acting director.
OpenBMB open-sources two models: MiniCPM-RobotManip (1.5B VLA for robotic manipulation) and MiniCPM-RobotTrack for tracking. The models enable robots to understand, remember, and act in physical environments.
Analysis from Ars Technica argues that the next frontier in AI-assisted development is not better models but better 'harnesses' that manage context, with examples including Augment Code and Claude Code. The piece interviews developers on moving beyond simple grep-like tools to context-aware coding agents.
Alibaba will ban employees from using Anthropic's Claude Code starting July 10, classifying it as high-risk software. Anthropic's Thariq Shihipar confirmed an experiment that secretly identified Chinese users, and Alibaba recommends its own Qoder tool instead.
Databricks blog post explains how to scale document classification to over 100,000 labels in production. Covers techniques for handling extreme multi-label classification at scale.
Robotix Sally, a silicone-skin humanoid robot, will teach AI to 11th and 12th graders in a New York school this autumn in a first-ever US experiment. The initiative aims to integrate AI and robotics into classroom learning.
Claude Code v2.1.181 (released June 17th) uses the Rust port of Bun, with startup 10% faster on Linux. The switch was described as "boring is good" and went mostly unnoticed by users.
Ishita Daga of Tesla argues that most enterprise agents fail because they lack understanding of business data structures. The fix is building semantic structure, not longer prompts or bigger models.
MCP's 2026-07-28 specification — its largest revision since launch — makes the protocol core stateless, enabling horizontal scaling, serverless deployments, and round-robin load balancing on standard HTTP infrastructure. Amazon Bedrock's AgentCore Gateway already supports the new spec, which is maintained under the Linux Foundation's Agentic AI Foundation.
Get tomorrow's AI brief in your inbox