AI News
Latest AI engineering news, updated daily.
Ai Agents
AWS OpenSearch and Cloudflare Mesh Pivot to Agent Workloads
AWS and Cloudflare have overhauled their core infrastructure to treat autonomous AI agents as first-class clients as machine traffic surges.
Autonomous Agents · Cloud Infrastructure · Machine Traffic
Ai Engineering
Tunix Hackathon Yields 1B-Parameter Gemma Reasoning Models
Google released the results of its Tunix hackathon, showcasing how developers trained small Gemma models to use reasoning traces on a strict compute budget.
Gemma Models · Reasoning Models · Fine Tuning
Prompt Engineering
Prompt-Driven Custom Feeds Bypass YouTube's Standard Algorithm
YouTube introduced conversational prompts to generate dynamic video feeds, alongside mandatory disclosure labels for photorealistic AI content.
Youtube Ai · Conversational Search · Custom Feeds
Ai Engineering
$300M SN50 Chip Order Validates SambaNova's ASIC-Native Cloud
General Compute has launched an inference neocloud with a $300 million order of air-cooled SambaNova SN50 chips capable of 700 tokens per second.
Ai Hardware · Sambanova Sn50 · Inference Cloud
Ai Agents
Parallel Search Powers Sesame's New iOS Voice Agent App
The Oculus founders' startup Sesame has launched a public preview iOS app featuring low-latency voice agents driven by simultaneous parallel search.
Conversational Ai · Parallel Search · Voice Technology
Ai Agents
Cloudflare Ships Skipper AI Agent and Town Lake Data Platform
Cloudflare launched Town Lake and the Skipper AI agent to consolidate massive internal data sprawl into a single SQL interface with natural language querying.
Cloudflare · Data Infrastructure · Natural Language Querying
Ai Agents
Task-Scoped Permissions Arrive in Anthropic Zero Trust
Anthropic released a technical framework for securing autonomous AI systems, introducing machine-verifiable identities and just-in-time access controls.
Zero Trust · Ai Security · Autonomous Agents
Prompt Engineering
Multi-Turn Attacks Erode Safety Guardrails in 15 AI Models
Cisco researchers reveal that multi-turn prompt attacks dramatically increase vulnerability success rates across 15 proprietary AI models, including GPT-5.4.
Ai Safety · Prompt Injection · Vulnerability Research
Ai Agents
CodeRabbit Routes Claude 4.x Models to Fix AI Intent Gaps
CodeRabbit’s new orchestration layer uses Claude Opus 4.7 and Sonnet 4.6 to translate high-level Jira requirements into validated coding plans before execution.
Anthropic Claude · Ai Orchestration · Automated Code Review
Ai Coding
Claude 4 Engineering Edition Solves 48.2% of SWE-bench 2026
Anthropic released Claude 4 Engineering Edition with a 2.5-million-token context window, autonomous IDE integration, and per-resolved-issue billing.
Anthropic Claude · Swe Bench · Autonomous Coding
Ai Engineering
Cascaded Speech Pipeline Brings Reachy Mini Inference Local
Hugging Face released an offline conversational stack for the Reachy Mini robot that replaces cloud APIs with a local pipeline built on Gemma 4 and Qwen3-TTS.
Robotics · Edge Computing · Offline Inference
Ai Agents
Frontier Agents Score Below 50% on SRE Task Benchmark
IBM Research and Artificial Analysis launched ITBench-AA, revealing that top frontier AI models score below 50% on complex enterprise SRE tasks.
Frontier Models · Site Reliability Engineering · Enterprise Ai
Ai Engineering
$8.2M Seed Backs Human Archive's Gig-Worker Robotics Dataset
Human Archive has raised $8.2 million to build a multimodal robotics dataset by paying Indian gig workers $1 per hour to record physical service tasks.
Robotics Data · Multimodal Datasets · Seed Funding
Ai Agents
Starlette BadHost Flaw Enables Auth Bypass in Python AI Agents
A critical HTTP Host header vulnerability in the Starlette framework allows attackers to bypass middleware authentication across the Python AI agent ecosystem.
Starlette Vulnerability · Python Security · Authentication Bypass
Ai Agents
Hugging Face Defines the Scaffold vs Harness Agent Architecture
Hugging Face has published a new technical glossary formalizing the structural differences between an AI agent's scaffolding and its execution harness.
Hugging Face · Agent Architecture · Technical Glossary
Ai Agents
Anthropic Moves Claude Mythos Toward Public Agent Access
Anthropic's autonomous vulnerability discovery model, Claude Mythos, has appeared in Claude Code, suggesting an upcoming public release for the restricted tier.
Anthropic · Claude Mythos · Autonomous Agents
Ai Agents
Android XR Launches With Gemini 3.5 Wearable Agent Support
Google's Android XR platform introduces a two-tier hardware strategy for smart glasses, relying on Gemini 3.5 to process multimodal agentic workflows.
Android Xr · Google Gemini · Multimodal Ai
Ai Engineering
DharmaOCR 7B Proves Domain Alignment Beats Parameter Scaling
Dharma-AI has released two specialized OCR models, demonstrating that targeted training history outpaces general-purpose frontier models on structured tasks.
Optical Character Recognition · Domain Alignment · Model Specialization
Ai Engineering
NVIDIA Nemotron-Labs-Diffusion Yields 6x TPF Over Qwen3-8B
NVIDIA has released the Nemotron-Labs-Diffusion model family, introducing a joint autoregressive and diffusion training objective to accelerate text generation.
Diffusion Models · Text Generation · Model Optimization
Ai Coding
Cursor Cloud Agents Shift to Isolated VMs and Durable Execution
Cursor has transitioned its AI agents to isolated cloud virtual machines with decoupled states and durable execution to handle multi-hour tasks.
Cloud Infrastructure · Durable Execution · Agentic Workflows