GPT-5.6 Sol Pricing Cut by 50%
Pricing cut 50% lets developers run complex reasoning tasks cheaper.
221 candidate events; 38 shortlisted stories; 31 with article text available; 12 selected. Counts describe different stages; clustering, validation and refills can change the shortlist.
Active filters: none. Sort: Ranked.
All 12 stories, in ranked orderPricing cut 50% lets developers run complex reasoning tasks cheaper.
Qwen3.8-27B enables on‑device coding and multimodal agents without cloud APIs, needing ~56 GB GPU memory at FP16.
CUDA Agent raises pass rate to 98.8% and speeds kernels 96.8% faster than torch.compile.
Nemotron 3.5 Lightning NVFP4 delivers up to 4× throughput while using only 22 GB memory after QAD compression.
A constraint‑aware allocator boosts GPU utilization by up to 33 percentage points versus FIFO scheduling.
Open-source DeepSeek Harness lets engineers build plugin-based AI agents, accelerating custom workflow development.
Hermes Bot Mode adds named bots with independent memory and skills, ready for desktop deployment under MIT license.
Anthropic’s Claude watermark embeds statistical patterns, but critics argue it may degrade text quality.
Google’s PhotoScan predicts insulin resistance from phone images with accuracy comparable to DXA scans.
HKU and KAI’s SMASH 2.0 robots completed a fully autonomous 11‑point table tennis game.
vLLM‑Omni’s Distributed Layerwise Offload runs a 124 GB model on 64 GB HBM, cutting host memory by 73%.
Cursor’s Origin offers a self‑hosted Git alternative, launched as GitHub suffered a prolonged outage.
No stories match. Active filters.