Exa Agent Ultra orchestrates subagent swarms for exhaustive deep research, scoring 81.4% on WANDR versus Opus 5.5's 72.3%.
Aikido Security has released Altar-1, its first open-weight security model. It is a compressed version of Z.AI’s GLM-5.3, ...
Perplexity Research pairs RFT with hint-guided self-distillation on GLM 5.2, reducing live tool-call failures by 21.2% ...
Black Forest Labs Releases FLUX 3 Action: A 7B Open-Weights World Action Model That Tops RoboLab-120
Black Forest Labs (BFL), the lab behind the FLUX image models, has released FLUX 3 Action. It is a 7B open-weights World ...
Fastino's GLiNER2.5-Decide is a 340M open-weight encoder that returns rule-constrained, scored decisions on CPU for routing and guardrails.
OpenAI launches GPT-6 Sol and Luna, lower-cost models priced from $0.10 per million input tokens with caching upgrades.
Google's Gemini 3.8 Flash TTS and Flash-Lite TTS add prompt-based voice design, 2,000+ voices, and 100+ language support.
Anthropic has released Claude Opus 5.5, the first model in its new Claude 5.5 family. The team states it performs at the ...
NVIDIA AI Releases SoL-Pi that uses auto-research loops to cut agent token traffic up to 49% and API cost about 33%.
BottleCap AI Releases ThinkingCap-Qwen3.8-27B: 37.2% Fewer Thinking Tokens at a 0.86pp Accuracy Cost
BottleCap AI's ThinkingCap-Qwen3.8-27B cuts thinking tokens 37.2% across 12 benchmarks with only 0.86pp accuracy loss. Drop-in for vLLM.
Contrastive-LM's CLM-8B scores agent actions instead of generating text, running up to 9× faster than TypeSafe's Jev zero-shot.
Voice input on phones has been solved for years. What has not been solved is the output. Speak into most dictation tools and ...
Some results have been hidden because they may be inaccessible to you
Show inaccessible results