
Hosted by Multiproduktion · EN

Google’s AI avatars for Workspace, 1Password’s Agentic Mode to let Claude log in securely, a survey exposing the agent evaluation gap and risky no-human deployments, and OpenAI’s GPT-Red automated red-teamer that uncovered “fake chain-of-thought” attacks—security implications.

Massive Suno leak shows YouTube scraping; a new enterprise report exposes the chatbot trap and orchestration risks. OpenAI encrypts internal agent communications, raising audit and safety concerns. GPT-Red auto-red-teams models - hardening defenses but creating dual-use hacking risks. What it means for creators, businesses, and security.

OpenAI’s rumored moving, screenless speaker aims to be a physical ChatGPT companion. Google quietly began using images, videos and voice to train Gemini—opt out in your Account settings. An OpenAI researcher eyes a potential $2B AI drug startup, and Spotify rolls out a ChatGPT-like assistant for Premium.

Richard Sutton launches Oak Lab to pursue continuous, self-supervised RL and challenge frozen LLMs. Stanford's TRACE turns agent failures into targeted training patches. Also: Nobel-backed warnings about AI's economic risks and Gemini bringing voice reports to Waze.

Breakdown of AgenticSTS's structured five-layer memory that shrinks prompts ~99% and wins at Slay the Spire 2; the rise of loop engineering/autoresearch for autonomous ML experiments; Anthropic's Claude 3.5 Sonnet 5x speedup in browser automation; and OpenAI's Canvas plus Advanced Voice live web search.

OpenAI launches GPT-5.6 (Sol/Terra/Luna) with programmatic tool calling; Meta rolls out Muse Spark 1.1 with a 1M-token context and active compaction; Anthropic's Jacobian lens peeks into Claude's internal reasoning; startup Lyzr used an agent to run a $100M raise. Big wins for agents — and new safety questions.

OpenAI’s GPT-Live enables real-time listen-and-talk conversations while GPT‑5.5 handles deep reasoning. General Intuition trains robots on video-game physics to build physical intuition. SpaceXAI’s Grok 4.5 is a fast, cheap coding/agent model. Claude Cowork outperforms Gemini on Gmail triage.

Anthropic's J-Lens reveals Claude's hidden 'J-Space' inner monologue, raising transparency and alignment concerns. Plus JadePuffer: a self-directed, adaptive AI ransomware; Liquid AI's Antidoom (FTPO) that fixes repeating doom-loops; and why open-source pressures frontier labs.

Google quietly started using uploaded media to train its AI—opt out in Web & App Activity > supplemental AI activity. AI model rankings are flipping weekly, Reddit is fighting AI-generated spam with LLMs, and JADEPUFFER reveals fully agentic ransomware. Check your settings and stay safe.

From Hollywood's 'don't-ask-don't-tell' use of AI deepfakes to search agents that fail to ask clarifying questions, plus LlamaIndex's legal-kb for precise document search and open-source PDF-to-JSON tools that unlock buried enterprise data.