Aikido Security has released Altar-1, its first open-weight security model. It is a compressed version of Z.AI’s GLM-5.3, ...
Perplexity Research pairs RFT with hint-guided self-distillation on GLM 5.2, reducing live tool-call failures by 21.2% ...
BottleCap AI's ThinkingCap-Qwen3.8-27B cuts thinking tokens 37.2% across 12 benchmarks with only 0.86pp accuracy loss. Drop-in for vLLM.
Black Forest Labs (BFL), the lab behind the FLUX image models, has released FLUX 3 Action. It is a 7B open-weights World ...
Fastino's GLiNER2.5-Decide is a 340M open-weight encoder that returns rule-constrained, scored decisions on CPU for routing and guardrails.
BottleCap AI has released ThinkingCap-Qwen3.8-27B, a fine-tune of Qwen3.8-27B that spends 37.2% fewer thinking tokens across 12 benchmarks. Macro accuracy moves from 86.65% to 85.79%, and long-context ...
Google's Gemini 3.8 Flash TTS and Flash-Lite TTS add prompt-based voice design, 2,000+ voices, and 100+ language support.
NVIDIA's Nemotron 3 Diarization is a 100M-parameter open-weight model that tracks 8 overlapping speakers in real-time streaming audio.
Nokia's open-source AnyJev turns open LLMs into calibrated decision models with no training, lifting automatable traffic 6.8x ...
Contrastive-LM's CLM-8B scores agent actions instead of generating text, running up to 9× faster than TypeSafe's Jev zero-shot.
Anthropic has released Claude Opus 5.5, the first model in its new Claude 5.5 family. The team states it performs at the ...