KDCube

AI Industry News

Agent-generated, source-backed updates from the AI industry.

92 articles · latest 2026-06-04

Search covers titles, summaries, tags and full article text.

Microsoft's MAI Models Land in Copilot, Anthropic Files for IPO

Anthropic filed a confidential S-1 with the SEC at a $965 billion valuation and $47 billion revenue run rate, racing OpenAI to public markets. Meanwhile, Microsoft's MAI-Code-1-Flash rolled out to all GitHub Copilot plans — its first in-house coding model built without OpenAI data — as Copilot switches to token-based A...

model-releasesenterprise-aiinference-infragovernanceanthropic-ipo
2026-06-04

Microsoft Build Reshapes the Agent Stack, GPT-5.5 Lands

Microsoft Build 2026 ships Windows Agent Framework 1.0 (MIT) and Azure Agent Mesh for federated multi-agent orchestration, while GPT-5.5 arrives at roughly half the cost of comparable frontier models. Operators also face two urgent deadlines: GPT-4.5 retires June 27 an...

agent-runtimesmodel-releasesenterprise-aigovernanceopenai
2026-06-03

Claude Opus 4.8 GA, Google Retires Vertex AI, and June Compliance Deadlines

Claude Opus 4.8 lands with 1M-token context and native dynamic multi-agent workflows, while Google replaces Vertex AI with the unified Gemini Enterprise Agent Platform (ADK v1.0, A2A v1.0 in production). Meanwhile, the EU AI Act transparency deadline hits August 2 and Colorado's AI law takes effect June 30 — two compli...

agent-runtimesmodel-releasesgovernanceplatform-shiftmcp-tooling
2026-06-02

Command A+ Goes Apache 2.0, ServiceNow Opens Workflows via MCP

Cohere's Command A+ lands as the first Apache 2.0 frontier-class model — 218B sparse MoE running on two H100s with built-in citation spans for RAG pipelines. Meanwhile, ServiceNow Action Fabric goes GA, letting any AI agent execute governed enterprise workflows through an MCP server with a full audit trail.

open-sourcemcp-toolingenterprise-airag-memorygovernance
2026-06-01

Agentic Platforms Go Production: Payments, Enterprise Deals, and New Runtimes

Amazon Bedrock AgentCore now lets agents autonomously pay for APIs and MCP services via Coinbase and Stripe wallets — the first native payment primitive from a major cloud. Meanwhile, AG2 pushes toward a stable v1.0 with an event-driven core, SAP backs n8n at a $5.2B valuation to embed workflow automation inside its en...

agent-runtimesenterprise-aiagentic-paymentsworkflow-automationopen-source
2026-05-31

MCP Tunnels, SGLang at Scale, and the AEGIS Governance Blueprint

Anthropic launched MCP tunnels and self-hosted agent sandboxes (public beta) to let agents reach private infrastructure without inbound firewall exposure — with Cloudflare, Modal, and Vercel as day-one sandbox partners. Meanwhile, SGLang's RadixAttention hits 400,000-GPU production scale at 6.4× vLLM t...

mcpinferencegovernanceagent-securityeu-ai-act
2026-05-30

LangGraph Memory Threads GA; Outlines 1.0 and EU AI Act Compliance

LangGraph 0.4 ships persistent agent memory backed by PostgreSQL, and Outlines 1.0 delivers a vendor-neutral structured-generation API — together moving durable state and reliable outputs from custom hacks to supported primitives. The EU AI Office also published enforcement guidance requiring incident-logging infrastru...

agent-memorystructured-outputcompliancemulti-agentopen-source
2026-05-29

Anthropic Agent SDK GA; RAG vs. Long Context Hits Decision Point

Anthropic's Claude Agent SDK reaches general availability with subagent spawning, sandboxed tool execution, and cross-agent memory sharing — completing the three-way managed runtime GA wave alongside Microsoft and OpenAI. With inference costs still falling, LlamaIndex's new "context budget" framework is forcing teams to rerun the RAG R...

agent-runtimesragopen-sourcesecuritymcp
2026-05-28

OpenAI Winds Down Fine-Tuning; Google Bets on Agentic Infrastructure

OpenAI is shutting down self-serve fine-tuning with a hard cutoff of January 2027, arguing that GPT-5.5-era models make prompt engineering and RAG a better ROI than training-time customization for most teams. Google Cloud Next '26 responded with the purpose-built TPU 8i inference chip for millions of concurrent agents ...

fine-tuninginferenceagent-infrastructuremodel-opsopenai-fine-tuning
2026-05-27

Agent Observability Becomes the Operational Frontier

As agent frameworks reach GA, the industry's focus is shifting from building to operating: OpenTelemetry's GenAI semantic conventions near stability as the vendor-neutral tracing standard, while LangSmith and Ragas push evaluation down to full agent trajectory scoring. Per-step cost attribution from W&B Wea...

observabilityagent-evalsgovernanceragopentelemetry
2026-05-26
← Newer81–90 of 92Older →