SevenTnewS

Annonces

L'actualité tech en une phrase — le fait le plus marquant de chaque article publié.

AI Agents24 juil. 2026

OpenCode reached $40M ARR and 4.6M weekly users after an Anthropic Claude ban drove developers to its open-source AI coding agent, CEO Jay V said on a YC podcas

LLMs & Models24 juil. 2026

Anthropic's Claude Opus 5 nearly matches its flagship Fable 5 on benchmarks at half the price but deliberately weakens cybersecurity.

AI24 juil. 2026

Microsoft's MAI models live in Bing and Office cut GPU costs 84-89% in production, altering AI cost math.

LLMs & Models24 juil. 2026

Alibaba's 0.8B OvisOCR2 end-to-end model beats every pipeline document parser with 96.58 on OmniDocBench.

Cybersecurity24 juil. 2026

Google's Gemini 3.5 Flash Cyber, a lightweight fine-tuned model locked to governments, found 55 V8 vulnerabilities vs Claude Opus 4.6's 36.

AI24 juil. 2026

Hugging Face revamps its Inference Endpoints service to cut deployment complexity, offering vLLM, SGLang, and custom containers with autoscaling and pay-as-you-

Benchmarks & Tests24 juil. 2026

OpenRTAG, a new benchmark from researchers, reveals how text-attributed graph models fail on dirty data, unifying nine degradation scenarios across nine dataset

AI24 juil. 2026

PRO-LONG, a July 2026 framework, uses programmatic memory to let LLM agents recall 47 steps back, boosting ARC-AGI-3 pass rates by 18 points.

Tools & Frameworks24 juil. 2026

A new arXiv paper repositions LangGraph as a workflow orchestration tool, not an AI benchmark, with three executable recipes for stateful processes.

AI24 juil. 2026

Voxtral TTS, a new hybrid text-to-speech model, beats ElevenLabs Flash v2.5 with a 68.4% win rate in voice cloning and releases its weights openly.

AI24 juil. 2026

New tutorial from researchers reveals production failures of LLM agents not captured by academic benchmarks, with mitigations including verification pipelines a

LLMs & Models24 juil. 2026

PoTRE, a new AI framework using four specialized reasoning agents, achieved 49.92% on Humanity's Last Exam by proving diversity of thought outperforms sheer mod

AI24 juil. 2026

Alibaba ATH's HappyOyster 1.0 builds a drivable 3D world from a single photo or text prompt, now in gray testing.

NLP & ML24 juil. 2026

Researchers artificially lesioned an AI model to reproduce stroke-induced aphasia errors, matching individual patient profiles with high accuracy.

Qwen / Alibaba24 juil. 2026

Qwen Cloud launched Token Plan Individual and Team tiers, bundling AI models to cut costs 40% and integrate with Claude Code and Cline.

AI24 juil. 2026

Researchers built a neural network that learns emotional associations via Pavlovian conditioning, reproducing human-like preferences for images.

Open Source24 juil. 2026

Google DeepMind's Gemma 4 crossed 300 million downloads, signaling open-weight models now outperform proprietary APIs on cost.

AI24 juil. 2026

July 22 arXiv study shows natural-language autoencoders pass reconstruction tests despite 98% false claims; RECAP achieves AUC 0.95 for verifiable explanations.

Tools & Frameworks24 juil. 2026

OpenCode adds inclusionAI's Ling 3.0 Flash for free, betting token efficiency wins the AI coding war.

AI24 juil. 2026

Sakana AI launched Fugu Ultra 1.1, a routing layer that hides which AI model processes your prompt via a single API to cut costs and complexity.

← PrécédentPage 1 / 3 · 59 annoncesSuivant →