Black Forest Labs (BFL), the lab behind the FLUX image models, has released FLUX 3 Action. It is a 7B open-weights World ...
Fastino's GLiNER2.5-Decide is a 340M open-weight encoder that returns rule-constrained, scored decisions on CPU for routing and guardrails.
BottleCap AI's ThinkingCap-Qwen3.8-27B cuts thinking tokens 37.2% across 12 benchmarks with only 0.86pp accuracy loss. Drop-in for vLLM.
OpenAI launches GPT-6 Sol and Luna, lower-cost models priced from $0.10 per million input tokens with caching upgrades.
Anthropic has released Claude Opus 5.5, the first model in its new Claude 5.5 family. The team states it performs at the ...
Google's Gemini 3.8 Flash TTS and Flash-Lite TTS add prompt-based voice design, 2,000+ voices, and 100+ language support.
Voice input on phones has been solved for years. What has not been solved is the output. Speak into most dictation tools and ...
Contrastive-LM's CLM-8B scores agent actions instead of generating text, running up to 9× faster than TypeSafe's Jev zero-shot.
SpaceXAI releases Grok Voice Transcribe 2.0, a speech-to-text API claiming 2x accuracy over 1.0 at $0.10 per hour.
NVIDIA's Nemotron 3 Diarization is a 100M-parameter open-weight model that tracks 8 overlapping speakers in real-time streaming audio.
Nokia's open-source AnyJev turns open LLMs into calibrated decision models with no training, lifting automatable traffic 6.8x ...
Kyutai's Voice of Reason uses reinforcement learning to lift GLM-4-Voice from 27.3% to 77.1% on spoken GSM8K math.