Live50 new in 24h/Ollama: Ollama v0.34.4
Preview · "(demo)" items are samples. How we verify
● Live wire

What just shipped in AI

33 official sources. Every 10 minutes. Straight from the source, unedited.

Updated · 32/33 sources responding · refreshes every 10 min

Just inNew releaseOllama v0.34.4 is outNew: server: fix intermittent "model not found" errors · server: apply structured outputs in a single pass on thinking models · app: avoid System Events for ChatGPT/Codex detectionOllamaJust inNew researchCOMED: The Missing Middle Between Routing and Collaboration in Multi-LLM InferenceNo single Large Language Model (LLM) is uniformly reliable across queries, motivating multi-model inference systems that either route among models or combine their outputs.arXiv cs.CLJust inNew researchExperts Rise Where LLMs Disagree: Using Cross-Model Disagreement to Target Expert Effort in LLM Codebook Revision for Large-Scale AnnotationLarge-scale text annotation brings expert insight to millions of documents, often through a codebook that AI annotators follow. Developing a robust codebook, however, takes months.arXiv cs.CLJust inNew researchThe Drift Contract: Spectral Updates for Depth-Robust Local LearningLocal learning trains each layer with its own auxiliary loss and no global backward pass, which makes layer updates structurally parallel.arXiv cs.LGJust inNew researchSignal2Symbol: Neuro-Symbolic Temporal Reasoning for Explainable Physiological Time-Series Anomaly DetectionPhysiological time series such as electrocardiograms (ECG) and electroencephalograms (EEG) exhibit complex temporal structure, substantial acquisition variability, and a strong need for transparent decision-making.arXiv cs.LGJust inNew researchDo Synthetic Personas Predict Real Audience Response? A Sim-to-Real Study Where a No-Persona Baseline Beats Persona-Based Copy SimulationMarketers increasingly use large language models (LLMs) as "synthetic personas" to predict how an audience will react to a piece of copy before it ships, encouraged by evidence that profile-conditioned LLMs mimic human samples.arXiv cs.AIJust inNew researchDo Existing Preconditioners Improve Biomedical Tabular Foundation Learning? An Empirical Study on TabPFN OptimizationTabular foundation models have recently shown strong potential for structured biomedical data analysis. Among them, TabPFN has emerged as an effective approach for low-data tabular classification tasks.arXiv cs.AIIn the newsEverything new coming to Meta’s AI agent MuseCEO Mark Zuckerberg kicked off the company’s annual Connect event in Menlo Park on Wednesday with a keynote that made one thing clear: Meta is going all-in on Muse. It's even coming to Meta's AI glasses.TechCrunch AIcoverageIn the newsMeta made a Tamagotchi-like wearable for its Muse AI agentThe tiny hardware device creates another mobile home for its AI agent Muse.TechCrunch AIcoverageNew releaseGemini CLI v0.61.0 is outNew: Changelog for v0.60.0-preview.0 · Changelog for v0.59.0 · preserve explicit versioned Flash model IDsGemini CLIIn the newsMeta introduces camera-free AI glassesMeta says the camera-free glasses will be lighter and have up to 12 hours battery life.TechCrunch AIcoverageNVIDIAIntroducing NV-Reason-CT Open 3D CT VLM for Radiologist Chain-of-Thought ReasoningRadiology AI has made remarkable strides in detecting abnormalities across chest X-rays, pathology slides, and 2D scans. Yet one of the most clinically rich andNVIDIA DeveloperIn the newsAnthropic says its biology lab has already found something bigBut maybe the biggest reveal is that Anthropic has not let Claude run loose in its biology lab. Humans are still, so far, in the loop.TechCrunch AIcoverageNew modelGLM 5.3 Prime is now available$2.8 in · $8.8 out per 1M tokens. 1M context. GLM-5.3-Prime is the high-speed variant of Z.ai's GLM-5.3, inheriting its full capabilities while delivering 1.5–2× the output throughput through inference acceleration.Z.aiNew releaseClaude Agent SDK v0.2.159 is outNew: Internal/Other Changes · Updated bundled Claude CLI to version 2.1.281 · Pinned default model for e2e tests to claude-opus-5 to work around CI failures with the CLI's new default modelClaude Agent SDKNVIDIAValidate GPU Cluster Readiness Before AI Workloads LandA GPU cluster can pass every health check and still fail to run an AI workload. Even when every GPU, network link, and pod reports healthy, a 512-GPU trainingNVIDIA DeveloperIn the newsEnveda secures $311M to bring more nature-derived AI drugs into clinical trialsThe round valued the AI biotech at $2 billion. It is currently testing drugs that treat skin conditions and preserve weight loss after stopping GLP-1s.TechCrunch AIcoverageNew modelQwen3.8 Max Prime is now available$4 in · $12 out per 1M tokens. 1M context. Qwen3.8 Max Prime is a higher-throughput variant of Qwen3.8 Max from Alibaba's Qwen team, served as a separate SKU at a higher price point. It accepts text, image, and video...QwenHugging FaceHow to Use NVIDIA Warp and MjWarp to Accelerate Robotics Simulation and Learning WorkflowsHugging FaceAWSFrom portal-hopping to instant answers: HEMA’s journey with MCP and Amazon BedrockThis post is co-written with Mauro Rallo and Patrick van der Plas from HEMA. When engineers at HEMA needed an answer, they went portal-hopping, navigating disconnected wikis, service catalogs, and IT portals to find it.AWS Machine Learning blogNVIDIAManage Kubernetes Node Fleets with NodeWrightKubernetes manages what runs on your nodes. Managing the nodes themselves is the challenge: kernel settings, system packages, storage layouts, security agents,NVIDIA DeveloperAWSAgentic conversational video intelligence built on AWSWith video intelligence powered by agentic AI, you can ask natural language questions about uploaded videos and get answers within seconds.AWS Machine Learning blogAWSUse open weight models as your AI coding agent with Amazon BedrockAI coding agents have become a core part of how developers write, debug, and refactor software. Open weight models on Amazon Bedrock now make these agents practical to run privately and cost-effectively.AWS Machine Learning blogHacker News buzzClaude discovers a novel enzyme system with CRISPR-like repeats559 points and 582 comments on Hacker News.Hacker NewscoverageNew releaseLangGraph 0.4.32 is outNew: Changes since cli==0.4.31 · place self-hosted deployments on a listener · clarify agent flags and support env defaultsLangGraphGoogleGoogle Beam expands with new regions, partners, and customersWe’re expanding Google Beam to five new countries, and partnering with Industrious for an extended network.Google AI blogIn the newsChatGPT Voice gets closer to "Her" with email, calendar, and Slack accessChatGPT Voice now runs on OpenAI's new GPT-6 Astra, Sol, and Luna models and can tap into plugins like email, calendar, and Slack. Users can manage appointments, send emails, or build websites just by talking.The DecodercoverageAWSAmazon Bedrock Managed Knowledge Base now supports Salesforce and Zendesk as native data source connectorsAWS announces Salesforce and Zendesk data source connectors for Amazon Bedrock Managed Knowledge Base, a fully managed retrieval-augmented generation (RAG) service.AWS What's NewIn the newsGoogle's new Flash TTS models let you design AI voices from scratch using text descriptionsGoogle is introducing two new text-to-speech models, Gemini 3.8 Flash TTS and Flash-Lite TTS, which support more than 100 languages.The DecodercoverageIn the newsGemini 3.8 TTS PlaygroundTool: Gemini 3.8 TTS Playground Google released two new Gemini text-to-speech models today - gemini-3.8-flash-tts and gemini-3.8-flash-lite-tts .Simon WillisoncoverageIn the newsChatGPT mobile app gets voice-based agentic featuresPro and Plus users will be able to use the Work tab on their phones to complete agentic tasks.TechCrunch AIcoverageIn the newsYouTube adds AI tools to Creator Studio with script coaching, smart thumbnails, and Gemini editingYouTube is adding AI tools to its creator studio. A storytelling assistant analyzes scripts and rough cuts, Gemini becomes a chat-based editing assistant for Shorts, and a new live translation feature turns English streams into…The DecodercoverageMicrosoftOffloaded inference for real-world physical AI roboticsAt a glance Challenges a core assumption in robotics AI: Our research shows that running physical AI inference exclusively on onboard GPUs can limit robot performance, battery life, and scalability, and that offloading inference…Microsoft ResearchGoogle DeepMindAdvancing Private AI Compute with secure, server-side memoryIntroducing private, server-side memory to Private AI Compute for personal AI.Google DeepMindOpenAITwo years of OpenAI AcademyMarking two years of OpenAI Academy and bringing AI skills to even more communities.OpenAINVIDIAHow SWE-Serve Exposes the Gap Between Local Tests and Live ServingAn AI coding agent’s patch can pass tests yet fail when the server loads a real model and handles requests. Evaluating changes to inference-serving softwareNVIDIA DeveloperGoogle DeepMindGemini 3.8 text-to-speech says helloGoogle DeepMindIn the newsAnthropic engineer explains why Claude's writing got worse although the model got smarterAnthropic employee Jackson Kernion explains why newer Claude models write so oddly. Optimizing for math, code, and technical explanations aimed at other AI models has created a style that sounds like "overly-dense info dumps" to…The DecodercoverageNew modelSpace Bunny Alpha is now availableFree to use. 1M context. Space Bunny Alpha is an anonymous large model with blazing-fast inference, strong coding capabilities and native multimodal input support.OpenRouterIn the newsNvidia-backed Nscale keeps its biggest customer, Bytedance, out of its IPO filingNscale, the Nvidia-backed AI cloud provider, leaves its most important customer, Bytedance, out of the main prospectus for its planned US IPO.The DecodercoverageNew modelAion 3.5 Mini is now available$0.7 in · $1.4 out per 1M tokens. 262K context. Aion 3.5 Mini is a multi-model roleplaying and storytelling system from AionLabs, built on the GLM family of models. It is the smaller, lower-cost sibling of Aion 3.5 and uses...AionLabsNew modelAion 3.5 is now available$3 in · $6 out per 1M tokens. 262K context. Aion 3.5 is a multi-model roleplaying and storytelling system from AionLabs, built on the GLM family of models.AionLabsOpenAIOpenAI extends cyber access to Ukraine for civilian defenseOpenAI is extending access to its Daybreak program to the Government of Ukraine to support the cyber defense of civilian infrastructure.OpenAIHacker News buzzClaude Code reads AGENTS.md only when telemetry is on [fixed]460 points and 261 comments on Hacker News.Hacker NewscoverageOpenAISam Altman’s remarks at the United Nations Security CouncilOpenAI CEO Sam Altman discusses AI safety, human control, and international cooperation in remarks to the United Nations Security Council.OpenAIOpenAIHarvey turns legal context into stronger drafts with GPT-6 AstraGPT-6 Astra produces more structured, context-aware legal documents, freeing lawyers to focus on strategy.OpenAIOpenAIHow invideo improves color grading 3x with GPT‑6 AstraWith GPT‑6 Astra, invideo plans edits with greater precision, improves color correction and grading threefold, and produces 50 custom effects in one day.OpenAIOpenAIRingg’s AI agents resolve up to 65% of customer calls with OpenAIUsing GPT-5.6, Ringg powers multilingual agents across voice, chat, WhatsApp, and web for 90% less cost vs. GPT-4.1.OpenAINew modelSolar Mini 4 is now available$0.05 in · $0.2 out per 1M tokens. 524K context. Solar Mini 4 is Upstage's compact, cost-efficient language model, a 35B-parameter mixture-of-experts with 3B active parameters and a 524K context window.UpstageOpenAIIntroducing MentalHealthBenchMentalHealthBench is an expert-informed benchmark for evaluating helpful and safe AI responses across realistic mental health conversations.OpenAIIn the newsSF October 14th: A Birds of a Feather Session on Agentic EngineeringSF October 14th: A Birds of a Feather Session on Agentic Engineering I'm hosting an evening event with Jesse Vincent in San Francisco on Wednesday 14th October for people who are building weird and interesting things with and on…Simon WillisoncoverageOpenAIChatGPT Ads expands to Southeast Asia and TaiwanChatGPT Ads is expanding to Southeast Asia and Taiwan, giving eligible businesses new ways to reach people across more than 60 countries.OpenAINew releaseClaude Agent SDK v0.2.158 is outverbatim_prompts option : Added ClaudeAgentOptions.verbatim_prompts (default False ). When True , user messages are delivered to the CLI exactly as written — no @path file expansion and no slash-command dispatch.Claude Agent SDKAWSAmazon CloudWatch Omni: AI-first observability for agents and applicationsAWS announces the general availability of Amazon CloudWatch Omni, an evolution of Amazon CloudWatch. Omni is an AI-powered observability experience organized around your teams and the applications they run, so that you can…AWS What's NewVercel AI GatewayGemini 3.8 text-to-speech models now available on AI GatewayGemini 3.8 Flash-Lite TTS and Gemini 3.8 Flash TTS from Google are now available on AI Gateway . Both models take text and generate speech in more than 100 languages.Vercel AI GatewayIn the newsClaude Opus 5.5, GPT-6 Sol, GPT-6 Luna, and a new price warYesterday was Grok 4.7 ( pelicans ) and MiMo v2.6 Flash/Pro ( more pelicans ). Today Anthropic released Claude Opus 5.5 , and around an hour later OpenAI released GPT-6 Sol and GPT-6 Luna .Simon WillisoncoverageNew releaseOllama v0.34.3 is outNew: GET /api/show now advertises each model's thinking controls and default: · Available in the CLI with: · ollama show gemma4OllamaHacker News buzzPentagon says overreliance on AI contributed to missile strike on Iran school899 points and 507 comments on Hacker News.Hacker NewscoverageNew modelCommand A+ is now available$0.3 in · $1.5 out per 1M tokens. 192K context. Command A+ is Cohere's flagship model for enterprise agentic workflows. It accepts text and image inputs with a 192K context window, supports native tool calling with strict tool…CohereNew modelGPT-6 Luna Pro is now available$0.1 in · $0.5 out per 1M tokens. 1.1M context. GPT-6 Luna Pro is the same underlying model as [GPT-6 Luna](https://openrouter.ai/openai/gpt-6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on…OpenAINew modelGPT-6 Luna is now available$0.1 in · $0.5 out per 1M tokens. 1.1M context. GPT-6 Luna is the fast, cost-efficient model in OpenAI's GPT-6 series, positioned below GPT-6 Sol.OpenAINew modelGPT-6 Sol Pro is now available$2 in · $10 out per 1M tokens. 1.1M context. GPT-6 Sol Pro is the same underlying model as [GPT-6 Sol](https://openrouter.ai/openai/gpt-6-sol), served with `reasoning.mode` set to `pro` for higher-quality responses on complex…OpenAINew modelGPT-6 Sol is now available$2 in · $10 out per 1M tokens. 1.1M context. GPT-6 Sol is the cost-efficient high-end model in OpenAI's GPT-6 series, positioned below the flagship GPT-6 Astra and above the fast GPT-6 Luna tier.OpenAIAWSBring more intelligence to everyday work with GPT-6 Sol and GPT-6 Luna on Amazon BedrockGPT-6 Sol and GPT-6 Luna are now generally available on Amazon Bedrock, giving you more options to match intelligence and efficiency to each workload.AWS Machine Learning blogAWSOpenAI GPT-6 Sol and GPT-6 Luna are now generally available on Amazon BedrockToday, AWS announces the general availability of GPT-6 Sol and GPT-6 Luna from OpenAI on Amazon Bedrock. Expanding the GPT-6 family alongside Astra, these two models give teams more ways to balance intelligence, speed, and cost…AWS What's NewHacker News buzzGPT-6 Sol and Luna1739 points and 826 comments on Hacker News.Hacker NewscoverageAWSClaude Opus 5.5 is now available on AWSToday, we’re excited to announce the availability of Claude Opus 5.5 on Amazon Bedrock and Claude Platform on AWS , the first of the Claude 5.5 model family.AWS Machine Learning blogNVIDIAEnabling Private High-Performance Production AI Inference with NVIDIA Confidential ComputingAs large language model (LLM) inference increasingly processes sensitive information and proprietary model context across personal, enterprise, and regulatedNVIDIA DeveloperAWSEvaluate skill-equipped agents with Strands Evals and Amazon Bedrock AgentCoreGeneral-purpose agents handle a broad range of tasks, but you still need them to follow the procedures that run your business: compliance checks, document-processing workflows, escalation policies, engineering conventions.AWS Machine Learning blogNVIDIATopology-Aware Workload Scheduling with NVIDIA TopographAI factories are power-limited systems that deliver maximum value when fully optimized. GPU workload placement is a key optimization. Poor workload placementNVIDIA DeveloperNew modelClaude Opus 5.5 is now available$4 in · $20 out per 1M tokens. 1M context. Claude Opus 5.5 is Anthropic's flagship model for demanding reasoning, coding, and long-horizon agentic work, succeeding Claude Opus 5.AnthropicHacker News buzzClaude Opus 5.51771 points and 1099 comments on Hacker News.Hacker NewscoverageNew releaseAnthropic Python SDK v1.8.0 is outapi: add support for claude-opus-5-5, inline tool definitions and MCP tool-list pinning (beta)Anthropic Python SDKAWSClaude Opus 5.5 is now available on AWS GovCloud (US)AWS GovCloud (US) now offers Claude Opus 5.5, Anthropic’s most capable Opus model yet and, the first of the Claude 5.5 model family, a better collaborator that handles long-running coding and knowledge work, reporting back…AWS What's NewHacker News buzzOpenAI GPT–6 Astra breaks Enigma message that has resisted solution since 2005726 points and 437 comments on Hacker News.Hacker NewscoverageAWSAWS Security Hub AI Inventory adds Azure self-hosted instance supportAWS Security Hub AI Inventory now supports discovering and cataloging AI assets running on self-hosted instances in Microsoft Azure.AWS What's NewNew releasevLLM v0.30.0 is outThis release features 762 commits from 315 contributors (104 new)!vLLMAWSAmazon Connect Customer launches agent-to-agent collaborationAmazon Connect Customer now supports agent-to-agent collaboration, giving customers the choice to bring in specialized AI agents during a live interaction to resolve a customer request.AWS What's NewHugging FaceHow UK AISI and EvalEval Are Making Benchmark Results ReproducibleHugging FaceHugging FaceTransformers now runs llama.cpp quantsHugging FaceHugging FaceJun Kim, oMLX creator and maintainer, joins Hugging Face to support the MLX communityHugging FaceVercel AI GatewayGPT-6 Sol and Luna now available on AI GatewayGPT-6 Sol and GPT-6 Luna from OpenAI are now available on AI Gateway . Both models bring GPT-6 improvements in professional work, coding, computer use, factuality, and communication at a lower price than GPT-6 Astra .Vercel AI GatewayVercel AI GatewayClaude Opus 5.5 now available on AI GatewayClaude Opus 5.5 from Anthropic is now available on AI Gateway . It is a step-change improvement over Opus 5 , with its biggest gains in agentic coding, long-running agent tasks, and knowledge work.Vercel AI GatewayNew modelMiMo-V2.6-Pro-UltraSpeed is now available$4.35 in · $8.7 out per 1M tokens. 1M context. MiMo-V2.6-Pro-UltraSpeed is the fast speed edition of Xiaomi's flagship foundation model, MiMo-V2.6-Pro.XiaomiNew modelMiMo-V2.6-Flash is now available$0.14 in · $0.28 out per 1M tokens. 1M context. MiMo-V2.6-Flash is an open-source foundation model developed by Xiaomi. Built on a Mixture-of-Experts architecture with 309B total parameters and 15B activated per token, it employs…XiaomiMicrosoftImproving synthesis prediction of small molecules at scale with RetroChimeraAt a glance We report on the recent publication of our retrosynthesis model RetroChimera in the journal Nature (opens in new tab) .Microsoft ResearchNew releaseLangGraph 1.2.12 is outNew: Changes since 1.2.11 · add response_schema to interrupt() · type undeclared v3 stream projectionsLangGraphHugging FacePruning LLMs Like a Physicist: Block Removal as an Ising Optimization ProblemHugging FaceHugging Facetokenizers v1: encode, decode and scaling, measuredHugging FaceVercel AI GatewayMiMo V2.6 models now available on AI GatewayMiMo V2.6 Pro , MiMo V2.6 Flash , and MiMo V2.6 Pro UltraSpeed from Xiaomi are now available on AI Gateway . MiMo V2.6 combines coding, reasoning, and tool use with native text, image, audio, and video understanding.Vercel AI GatewayNew releaseColibrì 1.12.0 is outNew: colibri 1.12.0 · 81 pull requests since v1.11.0. A new way to ask a model a closed question, · a redesigned dashboard and landing page, and a long run of small failuresColibrìAWSAWS Continuum now supports credential testing and accessible domain suggestionsAWS Continuum for penetration testing is a frontier agent that proactively secures applications throughout the development lifecycle by offering on-demand, customized penetration testing with real exploitability testing.AWS What's NewAWSAWS Resilience Hub adds three new capabilitiesThe next generation of AWS Resilience Hub is a central location in the AWS that helps platform engineering and site reliability teams assess and strengthen the resilience of their workloads running on AWS.AWS What's NewNew releaseAnthropic Python SDK v1.7.0 is outNew: api: add group with display_name to rate limits, deprecate group_type · tools: add compact_before_next_turn() to the tool runner · bedrock: raise an API error for eventstream exception and error framesAnthropic Python SDKGoogle CloudChanging the game: Using agentic AI to secure infrastructure codeAI is accelerating software development at an unprecedented pace. But as code generation scales, so do the challenges of securing the code, especially emerging AI-based vulnerability exploitations.Google Cloud AI & MLAWSKimi K3 by Moonshot AI is now generally available on Amazon BedrockAmazon Bedrock continues to expand its open weight model portfolio with the same security and governance that customers rely on.AWS What's NewGoogleNew experts join Google’s AI & Economy teamWe are expanding our AI & Economy team with world-class academic advisors, fellows, and core internal researchers.Google AI blogGoogleCo-creating the future of fashion with GoogleGoogle worked side-by-side with designers Jane Wade and Sergio Hudson to custom-design Google Flow tools to prep for NYFW.Google AI blogNew releaseOpenAI Agents SDK v0.22.3 is outNew: align conditional approvals with validated tool arguments · deliver tool-not-found output on server-managed resume · keep command paths POSIX on a Windows hostOpenAI Agents SDKGoogleMaking global data easier to exploreGoogle and the UN system have launched the UN System Data Commons, a new open platform making global statistics accessible and easy to search.Google AI blogNew releasevLLM v0.3.0 is outNew: Release vllm-proto 0.3.0vLLMNew releaseCrewAI 1.15.22 is outNew: Support aliases as connection identifiers · Record reasons for deployment creation failures · Collect human feedback and pause events in tracingCrewAIGoogle CloudHow Orange built FinOps accountability, and why agents are nextAt Orange , the leading France-based multinational telecom provider, there are days when engineering teams set aside their delivery backlogs and spend the day cleaning up cloud spend together. There's a leaderboard.Google Cloud AI & MLGoogle CloudCloud CISO Perspectives: How Google monitors AI threats and advances AI defensesWelcome to the first Cloud CISO Perspectives for September 2026. Today, Sandra Joyce shares the latest details on Google’s visibility into how attackers are using AI, and how we’re using AI to stop them.Google Cloud AI & MLMistralMistral and Mozilla are bringing open, private and multilingual AI to your web browserOpen, private and multilingual AI is coming to your web browser. Mistral and Mozilla team up to put powerful, trustworthy AI where you already browse.Mistral AIGoogle DeepMindIntroducing Gemini 3.8 Live and 3.8 Live Extended ThinkingGoogle DeepMindGoogleAI for Societal ImpactExplore this collection to see how experts and local leaders are using AI breakthroughs to ensure everyone can share the opportunity of AI.Google AI blogNew releaseColibrì v1.11.0 is outNew: 56 pull requests since v1.10.2. A ninth model family, five real bugs closed · across four engines, and the two platforms the C tests never built on now · building them in CIColibrìMicrosoftThe Economics of Agent Optimization: How AI agent governance controls cost and proves ROIThis blog post is the fourth and final installment of The Economics of Agent Optimization , which shares the strategies, capabilities, and proof points that can help you optimize agent costs and run AI as a managed investment…Microsoft Azure blogMicrosoftThe future of infrastructure resiliency starts with modernizationWhy infrastructure resiliency is essential for modern applications and AI workloads Organizations today face constant pressure to modernize; business-critical applications are being transformed, AI workloads are becoming…Microsoft Azure blogNew releaseTransformers 5.17.0 is outNew: Release v5.17.0 · New Model additions · Hy4-Preview is a 780B-parameter mixture-of-experts language model that activates 49B parameters perTransformersMistralCloudera and Mistral Partner to Bring Specialized, Sovereign Intelligence to Enterprise DataCloudera and Mistral join forces to bring specialized, sovereign AI intelligence to enterprise data, helping regulated industries innovate on their own terms.Mistral AINew releaseCrewAI 1.15.21 is outNew: Add telemetry to track checkpoint runtime and CLI usage · Fix gateway errors reported inside an HTTP 200 response · Keep deploy push on the AMP create sourceCrewAIGoogle CloudGoogle is a Leader in the 2026 Gartner® Magic Quadrant™ for Enterprise AI AssistantsWe are excited to share that Gartner has named Google a Leader in its inaugural 2026 Magic Quadrant for Enterprise AI Assistants .Google Cloud AI & MLNew releaseOpenAI Agents SDK v0.22.2 is outNew: support current image generation tool options · prevent UnixLocal file API symlink races · reset compaction response chain after popOpenAI Agents SDKMistralModernizing complex legacy code with AI agents.Mistral helped a European energy operator migrate 40,000 lines of Fortran 77 to C++. Learn how it was done, and the lessons to carry forward.Mistral AIMicrosoftBeyond the benchmark: How an adaptive approach drives scientific discoveryFor research and development (R&D) organizations, the promise of agentic AI is not a better one-time answer. It is a new way to explore complex scientific and engineering problems: pursuing multiple hypotheses, validating them…Microsoft Azure blogGoogle CloudHow KDDI built Buffmee, a faster, reliable consumer RAG appWhen building consumer-facing generative AI applications, balancing high generation quality with fast response times across diverse media types, can be challenging.Google Cloud AI & MLGoogle DeepMindAlphaGenome Atlas: A predictive map of every possible DNA letter change in the human genomeAlphaGenome Atlas maps the molecular effects of 9 billion single-letter DNA variants across the human genome.Google DeepMindMistralMistral raises €3B to make sovereign, open-weight AI the technology frontierMistral today announced that it has raised €3 billion in a Series D funding round at a post-money valuation of more than €21 billion.Mistral AIGoogle CloudSpanner migrations: Automating dual-write with Antigravity CLI for minimal disruptionWhen Google's Finance Engineering team needed to modernize their legacy data layer, they chose Spanner , a globally distributed, strongly consistent, multi-model database with high availability capabilities.Google Cloud AI & MLMicrosoftEnterprise AI transformation relies on the end-to-end platform: Azure was built for this momentSummary The recognition for Microsoft over the past couple of weeks comes down to models, infrastructure, data, applications, and developer tools working as one system when AI moves into production.Microsoft Azure blogMicrosoftGPT-6 Astra: Frontier intelligence for work, now generally available in Microsoft FoundryThe next era of enterprise AI will not be defined by chat experiences. It will be defined by how well a model can work for and with you.Microsoft Azure blogGoogle DeepMindIntroducing WeatherNext 3, our most advanced and accurate global weather AI modelGoogle DeepMindMicrosoftGigaPath-Flash and GigaTIME-Flash: Toward population-scale discovery with efficient pathology foundation modelsAt a glance The Flash family extends GigaPath and GigaTIME with dramatically improved efficiency, making large-scale pathology research more accessible and practical.Microsoft ResearchNew releaseTransformers v5.16.0 is outNew: Release v5.16.0 · New Model additions · Qwen4-Exp builds on Qwen3.5's hybrid text and multimodal architecture with three key components: GatedResidual (GR), Qwen Sparse Attention (QSA), and Per-Layer Embedding (PLE)TransformersMicrosoftMindTopo reveals VLMs’ spatial reasoning abilitiesAt a glance MindTopo is a new benchmark for testing topological reasoning in AI, evaluating whether multimodal models can understand concepts such as connectivity, enclosure, order, separation, and knots.Microsoft ResearchNew releaseModel Context Protocol 2026-07-28 is outThis release marks the stable release of the 2026-07-28 revision of the Model Context Protocol.Model Context ProtocolNew releaseModel Context Protocol 2025-11-25 is outThis release marks the stable release of the 2025-11-25 revision of the Model Context Protocol.Model Context Protocol
Source status
  • OpenRouter (14)
  • OpenAI (8)
  • Google DeepMind (5)
  • Google AI blog (5)
  • Mistral AI (4)
  • Microsoft Research (4)
  • Hugging Face (6)
  • NVIDIA Developer (6)
  • AWS What's New (10)
  • AWS Machine Learning blog (6)
  • Google Cloud AI & ML (6)
  • Microsoft Azure blog (5)
  • Anthropic Python SDK (2)
  • Claude Agent SDK (2)
  • OpenAI Agents SDK (2)
  • vLLM (2)
  • Ollama (2)
  • LangGraph (2)
  • Transformers (2)
  • PyTorch (0)
  • Model Context Protocol (2)
  • CrewAI (2)
  • Gemini CLI (1)
  • Colibrì (2)
  • arXiv cs.CL (2)
  • arXiv cs.LG (2)
  • arXiv cs.AI (2)
  • Vercel AI Gateway (4)
  • MarkTechPost (HTTP 403)
  • The Decoder (5)
  • Simon Willison (3)
  • Hacker News (6)
  • TechCrunch AI (6)