SevenTnewS

DeepSeek

DeepSeek-V3, DeepSeek-R1, Coder and the Chinese breakthrough in open source LLMs.

7 published articles

4 min read

AI Agents · Open Source

DeepSeek ships an agent harness where even the model is a plugin

DeepSeek released DeepSeek Harness, an open-source agent runtime where models, tools, sandboxes, and the UI are all Cordis plugins. Append-only session logs and a two-tool minimal mode point to a quieter ambition: auditable, reproducible agent runs.

2026-08-16

2 min read

Artificial Intelligence

DeepSeek-OCR 2 Brings Visual Causal Flow to Open-Source Document Understanding

DeepSeek-OCR 2 introduces Visual Causal Flow for human-like visual encoding. Now open-source on GitHub, the model supports vLLM and Transformers, dynamic resolution (up to 1,216 visual tokens), and document-to-markdown conversion.

2026-07-09

Featured2 min read

API economics

DeepSeek V4 drops in July with a pricing shift that will hurt if you code when the sun is up

DeepSeek V4 arrives mid-July with peak-load pricing. The API costs double during business hours, up to $1.74 per million output tokens, but off-peak rates can be a steal for cache-hit workloads.

2026-07-09

5 min read

Model Release

DeepSeek-V4 preview lands, and the open-weight math gets harder for everyone

DeepSeek unveils DeepSeek-V4 preview with top-tier reasoning and stronger agent abilities, now available across platforms. The release signals a major escalation in the open-weight AI race, putting pressure on OpenAI and Google to justify their closed-source strategies.

2026-07-05

2 min read

LLM Inference Optimization

Aleph Alpha builds a theoretical inference model to decode DeepSeek V3 performance from hardware primitives

Aleph Alpha's theoretical model predicts DeepSeek V3 inference performance from hardware parameters alone, revealing how GPU count and interconnect bandwidth shift the bottleneck between compute, memory, and communication.

2026-07-04

3 min read

Artificial Intelligence

DeepSeek's new model makes efficiency the AI arms race front line

DeepSeek has released a new LLM that sharpens its competition with major AI labs by betting on efficiency over brute force. The model aims to offer strong performance with optimized resource usage, reflecting a broader industry pivot toward cost-effective AI.

2026-06-30

4 min read

Open-source AI research

Million-token context on a tenth of the KV cache: DeepSeek-V4's efficiency bet

DeepSeek's V4 preview pairs a 1.6T-parameter Pro model with a 284B Flash variant, both at one million tokens of context. The paper claims 27% of the inference FLOPs and 10% of the KV cache of V3.2, the numbers that make the context cost-effective to serve.

2026-06-22