← Explore

Posts tagged with llm

Neural Dispatch · ·4 min read

Per-Token Costs Dropped 95%. Your AI Bill Went Up Anyway.

Gartner published a prediction on Sunday that deserves more attention than it's getting: inference costs per agentic workflow will increase more than...

inference-economicsagentic-aigartner
Neural Dispatch · ·5 min read

Eight Trillion Tokens Broke DeepSeek's Pricing Model

On August 1, DeepSeek's API processed 8 trillion tokens in a single day — 5 trillion from free-tier users alone.

deepseekapi-pricingself-hosting
Neural Dispatch · ·4 min read

Qwen 3.8 Max Goes Open Weight This Week — But "Open" Needs an Asterisk

Alibaba is about to release the weights for a 2.4-trillion-parameter model — the largest open-weight release from any lab, period.

qwenalibabaopen-weights
Data Eng Daily · ·5 min read

Flink Agents Puts an LLM Inside Your Checkpoint Boundary

Flink's checkpoint has always been a contract: snapshot the state, replay the log, get the same output. Apache Flink Agents 0.

flink-agentsstreamingllm
Neural Dispatch · ·5 min read

GPT-5.6 Goes Live: Three Tiers, Cerebras Speed, and a Model That Lies

OpenAI flipped the switch on GPT-5.6 yesterday.

gpt-5.6openaiai-safety
Neural Dispatch · ·5 min read

JADEPUFFER: An LLM Just Ran a Full Ransomware Attack, Debugging Included

Sysdig just published the first documented case of ransomware run entirely by an LLM agent. No human at the keyboard.

jadepufferransomwareai-security
Neural Dispatch · ·4 min read

Z.ai's GLM-5.1 Topped SWE-Bench Pro Without a Single NVIDIA Chip

An open-source model just claimed the top spot on SWE-Bench Pro — the benchmark that's become the de facto measuring stick for agentic software engineering.

glm-5.1z-aiopen-source
Data Eng Daily · ·5 min read

Vectorless RAG Is Real — But You Probably Shouldn't Throw Away pgvector Yet

PageIndex hit 98.7% accuracy on FinanceBench.

ragvector-databasepageindex