Explore how proactive AI agents from Meta, OpenAI and Uber shift technology from user queries to timely interruptions.
NVIDIA announced a new 64GB configuration of DGX Spark — from Acer, ASUS, Dell, Gigabyte, HP and MSI — its GB10-powered ...
Datalab's OmniExtractBench tests extraction on 620 documents from 4 vendor benchmarks with 1 deterministic scorer explaining ...
IBM makes Bob self-hostable: enterprises can run agentic software development on premises, in sovereign clouds, or fully ...
Prime Intellect launches Prime Inference, serverless and reserved serving for open models, running GLM-5.3 on NVIDIA GB200 ...
Microsoft AI's MAI-Transcribe-2-Streaming ranks #1 on Artificial Analysis with 2.5% WER at 0.13s, 60 languages, $0.54 per hour.
Decision AI models answer typed questions with calibrated probabilities instead of generated text. TypeSafe's Jev costs ...
IBM has made a self-hosted deployment of IBM Bob, its agentic software development platform, generally available. Enterprises can now run Bob on premises, in private or sovereign clouds, and in ...
The API model IDs are embed-v5.0-pro and embed-v5.0-fast, per Cohere’s model docs. Both output 2048, 1536, 1024, 768, 512, or 256 dimensions, with 2048 as default. Embeddings come back as float, int8, ...
Cloudflare releases Clef and Clef-flash, open-weight Apache 2.0 decision models returning typed probabilities for fast agent ...
AWS Strands Labs releases Strands Decider 2B, an open source decision model. It does not generate text. It reads a state and ...
Google DeepMind's Gemini 4 Argon brings 1M output tokens, leads DeepSWE v1.1, and launches at $2/$10 introductory pricing.