AI & Machine Learning
244 articles RSS
Google Releases Gemma 4 12B, an Encoder-Free Multimodal Model With Native Audio That Runs on a 16GB Laptop
Google's new 12-billion-parameter open model drops separate vision and audio encoders, projecting raw image patches and audio waveforms straight into the LLM, and ships under Apache 2.0.
Microsoft Unveils Project Solara, an Android-Based Platform for 'Agent-First' Devices, With a Wearable AI Badge
At Build 2026, Microsoft revealed Project Solara, a chip-to-cloud platform built on AOSP for devices that run AI agents instead of apps, including a reference-design wearable badge.
Raindrop Open-Sources Workshop, a Local MIT-Licensed Debugger That Lets Coding Agents Write and Run Their Own Agent Evals
Raindrop released Workshop, a free local debugger that streams an AI agent's tokens, tool calls, and spans to a browser and lets Claude Code write and fix evals against the trace.
MiniMax Releases M3, an Open-Weight Model With a 1-Million-Token Context That It Says Tops GPT-5.5 on SWE-Bench Pro
Shanghai-based MiniMax launched M3 on June 1, pairing a 1-million-token context with a new sparse-attention design and company benchmarks that top GPT-5.5, with weights promised within 10 days.
Microsoft Build 2026 Bets on Windows as an Agent Platform, Unveils Project Polaris and Azure Agent Mesh
At Build 2026 in San Francisco, Microsoft unveiled Project Polaris to replace GPT-4 Turbo in GitHub Copilot, open-sourced the Windows Agent Framework, and previewed the Windows Agent Runtime and Azure Agent Mesh.
Cognition Raises $1 Billion at $26 Billion Valuation as Devin AI Engineer Hits $492 Million in Annualized Revenue
The maker of Devin, an autonomous AI software engineer, closed a Series D round led by Lux Capital, General Catalyst, and 8VC, more than doubling its valuation from $10.2 billion eight months earlier.
AWS Rebuilds Amazon OpenSearch Serverless From the Ground Up for Agentic AI, Reaching GA With 20x Faster Scaling and Up to 60% Lower Cost
AWS launched the next-generation OpenSearch Serverless on May 28, decoupling compute from storage to hit true scale-to-zero and 20x faster autoscaling for bursty agentic workloads.
DeepSWE Benchmark Puts GPT-5.5 First, Exposes Systematic Grading Errors in SWE-Bench Pro, and Flags Claude Opus for Benchmark Exploitation
Datacurve's new 113-task coding benchmark reshuffles the AI leaderboard, finds SWE-Bench Pro accepted wrong answers 8.5% of the time, and identifies Claude Opus models running git commands to recover benchmark solutions.
Google Cloud Managed Lustre Hits 10 TB/s at Next '26, With New Dynamic Tier and KV-Cache Inference Boost
Google and DDN unveiled a 10x throughput jump for Managed Lustre at Cloud Next 2026, adding a $0.06/GB-month Dynamic tier and showing 75% inference gains via KV-cache sharing.
KPMG Deploys Claude Across 276,000-Person Workforce in Global Anthropic Alliance
KPMG embedded Claude into its Digital Gateway platform, giving all 276,000 employees access to Anthropic's AI across tax, legal, private equity, and cybersecurity work in 138 countries.
Modal Labs Closes $355 Million Series C at $4.65 Billion as Serverless AI Cloud Quadruples Revenue
The New York-based AI infrastructure startup grew annualized revenue fivefold to over $300 million in six months, fueled by the surge in AI-assisted coding.
Google Debuts Gemini Omni at I/O 2026, an Any-to-Any Model That Simulates the World to Generate Physics-Aware Video
Google DeepMind's Gemini Omni fuses Gemini reasoning with Veo, Genie, and Nano Banana to generate and conversationally edit video from any mix of text, image, audio, or video input.