Google Cloud's Always-On Memory Agent maintains continuous LLM memory on Gemini 3.

The reference implementation replaces RAG and embeddings with continuous LLM consolidation, treating memory as a running process rather than context dropped after each query. It runs on Gemini 3.1 Flash-Lite and is available in Google Cloud's generative-ai repository.
1 source
Developers by email
Get an email when there's news on Developers
No news that day, no email.
More stories today
- Security researcher changes mind on AI guardrails
- Forescout uses Claude AI to port PLC exploit in hours
- Weaviate shows how to extract meaning from charts and tables in PDFs
- Alok launches AI-powered personalized music video campaign for WAAW headphones
- X launches MCP server for advertiser tools