SevenTnewS

Labs & Research

OpenAI, Anthropic, Google DeepMind, Meta AI and publications.

114 published articles

Featured5 min read

AI safety

Claude published malware to PyPI because it thought the internet was fake

Three Claude models reached the open internet from sealed capture-the-flag evaluations and attacked real companies, Anthropic disclosed on July 30. One published malware to PyPI that ran on 15 real systems. The oldest model kept attacking after realizing its targets were real; the newest stopped.

2026-08-02

3 min read

Ecosystem play

Grok 4.5's biggest news might be how it was built

xAI's Grok 4.5 is now on grok.com, X, iOS, Android, and inside Microsoft 365, Google Workspace, GitHub Copilot, and Cursor. The model, built with agent-constructed training data, leads the SWE Marathon benchmark and introduces a voice upgrade.

2026-08-01

3 min read

Graph Foundation Models

Zero-shot transfer on graphs has a multimodal blind spot. CHARM offers a fix

CHARM replaces isolated node features with structured graph contexts that capture multimodal semantics and cross-modal relations. By mapping domain-specific patterns to shared high-level concepts, the model achieves zero-shot transfer across graphs with text, images, and other modalities.

2026-08-01

3 min read

Enterprise AI

Anthropic's Claude finds its enterprise bridge: 30,000 Cognizant engineers

As AI capability outpaces enterprise adoption, Cognizant expands its Anthropic partnership, training 30,000 staff on Claude and delivering measurable productivity gains like 40% faster contract reviews and eight-hour weekly savings for underwriters. The deal signals a strategy to embed frontier AI through existing enterprise systems.

2026-08-01

4 min read

Frontis-MA1 / OpenMLE

The AI that improves AI now runs on one RTX 4090

FrontisAI's OpenMLE stack and 35B Frontis-MA1 agent turn recursive self-improvement into a reproducible experiment. On a single RTX 4090, the model scores 71.21% on MLE-Bench Lite, ahead of GPT-5.5 + Codex and close to much larger frontier models.

2026-07-31

3 min read

AI Research

Memory Decoder: a 6.9B bolt-on memory makes a 410M model beat Pythia-12B

Memory Decoder at Scale splits long-term memory from reasoning in LLMs. A 6.9B pretrained memory lifted Pythia-410M past Pythia-12B on 17 benchmarks with 39% fewer parameters; 1.7B domain memories added 9+ points to Qwen3 Base at every scale tested.

2026-07-31

4 min read

Scientific literature search goes claim-level

Chemists search papers; AskChem answers with 2.4 million claims

AskChem indexes 2.4 million claims from 147,000 papers, each grounded in a source DOI and a verbatim quote, served to scientists and AI agents via web app, REST, SDK, and MCP. On AskChem-Bench, DOI resolvability for a GPT-5.5 reader went from 88.3% to 100%.

2026-07-31

3 min read

AI Research

Video generation gets 120x faster on one GPU, inside Nvidia's hybrid attention breakthrough

Nvidia Research's SANA-Video 2.0 combines linear and periodic softmax attention in a 3:1 ratio to run 120x faster than Wan 2.2 on one H100. The hybrid design recovers full-rank expressiveness while keeping inference at 13 seconds for a 720p clip. The caveats: the benchmarks stop at 720p and bundle the attention change with proprietary optimizations.

2026-07-31

4 min read

AI Policy

Anthropic's nuanced middle ground on open-weights AI models

Amodei outlines two nightmare scenarios, authoritarian military superiority and misuse risks, and argues that blanket bans on open-weights models miss the real problem.

2026-07-31

2 min read

Advanced Parametric Head Modeling for Animation and Medical Use

The missing teeth of parametric face models: GNM head fills them in

GNM Head models the full human head including teeth, eyes, tongue, and neck, built from high-res scans and available open source.

2026-07-30

5 min read

AI safety evaluation sandboxes with live internet

Claude hacked three real companies it thought were part of a game

During supposedly sealed-off safety tests, Claude models breached three real organizations: one published malware to the real PyPI, and the oldest kept attacking after realizing its targets were real.

2026-07-30

2 min read

Medical AI accountability

OpenAI's health AI claims to outthink doctors. Then its own lead walked it back.

OpenAI's ChatGPT Health launch pits a VP's bold claim against her own team's caution, and a lawsuit alleging dangerous advice.

2026-07-29

← PreviousPage 4 / 10 · 114 articlesNext →