The AI Index

235 entries20 categories172 people20 countriesUpdated 18 Sept

Latest

  1. PolicyTrump says he is forming an AI Force, and did not say what it would doNo details
  2. SafetyAnthropic put an outside evaluator inside the building, and it is a consultancy$1bn each
  3. PolicyCalifornia is studying whether to require a kill switch and auditors inside the labs16 Nov
  4. ModelsStepFun's 600B preview scores 44 and costs a dollar$1 / $2.70
  5. ModelsAlibaba's omnimodal flash model takes audio and video into a million tokens of context1M
10 more ↓Show fewer ↑
  1. LawUniversal and Sony sued Suno a second time, over the model it licensed60,202
  2. MoneyAnthropic's revenue run rate is heading past $100bn, and its listing is back on$100bn
  3. SafetyOne bug was in all four AI coding agents, because they all trusted the same thing4 of 4
  4. ModelsAnthropic says Claude now leads a quarter of the work that builds the next Claude26%
  5. ModelsGoogle's new voice model keeps talking to you while it goes off and does the work82.6
  6. PolicyThe web got a way to say no to AI training without disappearing from search17%
  7. PolicyThe US Treasury Secretary said the labs asking for safety rules are also asking not to be suedLiability
  8. ModelsApple shipped the rebuilt Siri, and said out loud that Google helped build the modelsEnglish
  9. ModelsA Chinese lab put a 744 billion parameter agent on the internet under the MIT licence, with no price and no blog post744B
  10. SiliconA network that does arithmetic on the way past raised $205m$205m

All the news, newest first →

The report

Your job did not disappear. It moved.

Almost nobody's work is being deleted. It is being renamed, split in two, or quietly stopped at the entrance. Here is where it went, and what the new version is called.

The most useful thing to understand about this moment is also the least dramatic: the overwhelming majority of jobs touched by AI are not being abolished. They are becoming a next version of themselves. The recruiter becomes the person who can tell a real machine learning engineer from a fluent one. The sales engineer becomes the person who demonstrates a system that is sometimes wrong. The lawyer becomes the person who audits a model's reasoning. The pattern repeats across almost every function, which means the useful question is not whether your job survives but what it is called now.

The second thing is harder to hear. The damage is real and it is concentrated almost entirely at the entrance. Employment for software developers aged 22 to 25 has fallen about a fifth since 2024, and the mechanism is not redundancy. Nobody was fired. Companies simply stopped opening the door, which is precisely why there was never a headline about it and why applying feels like shouting into an empty room. Against that, engineering hiring overall is down far less than the rest of tech. The profession is not shrinking. Its entrance is crowded.

The third is the one that pays. Some jobs are hiring hard while almost nobody is qualified for them, and they sit in unglamorous places: licensed electricians, liquid cooling technicians, commissioning agents, grid interconnection specialists, AI auditors. Several want a trade or a certificate rather than a degree. They are wide open for a dull structural reason, which is that the training takes four years and the demand arrived last year.

Search the fashionable title and you will find nothing, while the work is being hired for hard under a different name.
WasIs nowRecruiterTechnical recruiter, AISalespersonAI sales engineerSEO managerAnswer engine strategistCoderAI assisted engineerLawyerAI counsel and auditElectricianData centre electrician
Six of them. The same move repeats across nearly every function.

The comparison

Is my data safe with ChatGPT, Claude and Gemini?

Somebody in your office pasted an invoice into ChatGPT this week. Was that fine? It depends on the account they used, and almost nothing else.

You payThey takeFreethey use everything$20 a monthstill used, unless you switch it offA company accountnever used, and it is in writing

These companies are not charities and running this costs real money. So somebody pays. If it is not your credit card, it is your documents. On a free account, everything you send is used to train their AI: what you type, the files you upload, the photos, the voice notes. Paying twenty dollars a month changes almost nothing, because the vendors put the free plan and the paid personal plan in the same line of their own terms. It buys you features. It does not buy you a different contract.

A company account does. There, not using your documents is written into the agreement you sign rather than offered as a setting somebody has to remember. Of sixteen tiers checked against the vendors' own policies, every personal one that could be verified trains on your data, and not one company account does. So the fix is not finding a safer vendor. It is moving your staff off personal logins, and it costs about what they are already paying out of pocket.

The more you pay, the less they use it. That one sentence explains every row in the tables.

The guide

Can I put this into AI?

The question every organisation is stuck on, answered without the usual deflection about how you already use Gmail.

1Can this information go in?2Which platform may I use?3Can I rely on what came out?

No, you cannot hand it anything blindly. No, it is not all dangerous either. The same model is safe or unsafe depending on the door you walk it through: on your organisation's own account your text is not used to train anybody's model, it is held on your terms and it is auditable, and on a personal login none of that is true.

Which makes the decision smaller than it feels. What kind of information is this, which account am I using, and what is the thing allowed to do once it has it.

The discipline

How agents are actually built

PlanActObserveCheckthe check is the job

Nine decisions determine whether an agent works in production, and most teams get stuck on the second one because they collapse state and memory into the same idea. They are not the same idea. One is what the system knows right now, the other is what it is allowed to carry between runs, and confusing them is why agents behave well in a demo and badly on a Tuesday.

Read it9 phases, 9 worked examples

Or jump toState versus memory

Every box we have written →

The index underneath

Every model, chip, platform and lab we track, with specs, prices, licences and the people who built them. This is the material the writing above is drawn from, and every figure carries its source and its date.

01

How the field fits together

1ModelsThe frontier and the open weights.Opus 5 · GPT-5.6 · Kimi K3
2MediaImage, video, voice and music.Kling 3.0 · GPT Image 2
3ToolingAgents, routers and frameworks.Claude Code · LangGraph
4SiliconChips, clouds and capital.B300 · MI400 · RunPod

Each layer depends on the one beneath it. Silicon decides what can be trained, tooling decides what can be shipped, and the models everyone argues about sit on top.

02

The frontier right now

All 50
ModelContext$/Mtok out
GPT-6 AstraOpenAI1.05M$50
Claude Fable 5.1Anthropic1M$50
Claude Fable 5Anthropic1M$50
Claude Opus 5Anthropic1M$25
GPT-5.6 SolOpenAI1M+
Gemini 3.1 ProGoogle1M

Ranked by capability and adoption. A dash means the lab has not published a price.

03

Changed recently

  • NewQwen3.8-Omni-Flashreleased, text, image, audio and video into 1M tokens, API only09/18
  • NewStep 5 Previewreleased, 600B sparse MoE at $1 and $2.70 per million tokens09/18
  • NewGemini 3.8 Livereleased, first on the Artificial Analysis speech to speech index at 82.609/15
  • UpdatedSiri AIships on iOS 27, built with Google and its Gemini models09/14
  • NewAtria Dawn Previewreleased, 744B agentic weights under the MIT licence09/14
  • NewSuno v6released, first Suno model trained on licensed catalogues09/09

Releases, deprecations and policy moves. Every entry on this site carries the date it was last verified.

04

What a million tokens costs

Compare
  • GPT-6 AstraOpenAI$50
  • Claude Fable 5.1Anthropic$50
  • Claude Fable 5Anthropic$50
  • GPT-5.5OpenAI$30
  • Grok 4.6xAI$6
  • Grok 4.5xAI$6
  • Gemini 3.8 LiveGoogle$4.5

Output tokens, per million, and a 20× spread across the frontier. Capped at two models per lab; only 14 labs publish per-token prices at all, which is itself worth knowing.

05

Context windows

All models
  • Qwen3.8-Flash-NextAlibaba262K native, ~1M extended
  • Claude Sonnet 5Anthropic1M
  • GPT-5.6 LunaOpenAI1M+
  • InklingThinking Machines Lab1M
  • Atria Dawn PreviewShanghai AI Laboratory256K
  • Gemini 3.8 Live Extended ThinkingGoogleNot published

Sampled across the full range, largest to smallest. A window is a budget, not a promise. Recall falls as it fills, long before the limit.

06

The silicon underneath

All 8
ChipMemoryPrice
B300 Blackwell UltraNVIDIA288GB~$40–50K
B200NVIDIA192GB$30–50K
MI400AMD432GB
HAscend 950DTHuawei144GB HiZQ 2.0Above ¥250,000 (about $37K), reported
H100NVIDIA80GB$25–40K

Memory, not raw compute, is what usually decides which models a chip can serve.

07

How agents are built

The discipline
CONTEXTWINDOWPlanActObserveEvaluate

Nine decisions decide whether an agent works: scoping, state, memory, context, tools, guardrails, evals, observability, optimisation.

08

Who builds it

All 172
Isaac AsimovRen ZhengfeiMark ZuckerbergEmad MostaqueJeremy HowardSeymour PapertMax TegmarkJaron LanierStanisław LemMasayoshi SonElon MuskYoshua BengioJohn McCarthyJürgen SchmidhuberMelanie MitchellDaniel KahnemanZhang YimingYoav ShohamGreg BrockmanIan GoodfellowAllen NewellTimnit GebruShoshana ZuboffCasey NewtonLu QiLisa SuAlexandr WangAndrew NgJoseph WeizenbaumYuval Noah HarariBrian ChristianTed ChiangArvind KrishnaDemis HassabisGeoffrey HintonAlan Turing

172 people across 20 countries: founders, researchers, teachers, pioneers and the critics. This set rotates each week.

09

Where the money is

All 6
  1. 1MGXAbu Dhabi
  2. 2Thrive CapitalUnited States
  3. 3Andreessen HorowitzUnited States
  4. 4LightspeedUnited States
  5. 5NVIDIA / NVenturesUnited States

Sovereign funds now sit alongside the venture firms. The largest cheques in the field are no longer written on Sand Hill Road.

10

Trending on Hugging Face

All 50
  1. 1Qwen3.8-Flash-NextQwen208K
  2. 2GLM-5.3-Flashzai-org441K
  3. 3GLM-5.3zai-org94K
  4. 4Qwen3.8-Flash-Next-GGUFunsloth431K
  5. 5Qwen3.8-27BQwen5.0M

By Hugging Face's own trending score, which weights recent downloads and likes.

11

Top AI repositories

All 100
  1. 1ECCaffaan-m246K
  2. 2hermes-agentNousResearch240K
  3. 3deepseek-harnessdeepseek-ai208K
  4. 4tensorflowtensorflow198K
  5. 5AutoGPTSignificant-Gravitas187K

Stars, across fourteen AI topics. Stars measure attention rather than quality, but they measure it consistently.

12

What the field is arguing about

All 100
  1. 1AI;DR (AI; Didn't Read)1119
  2. 2Israel creates fake think tank in likely attempt to dupe AI chatbots1071
  3. 3Claude Fable 5.1 and Claude Mythos 5.11068
  4. 4Don't paste the AI, please1061
  5. 5GUIs should be fully keyboard-driven1048

Hacker News points over the last three weeks. What the field argues about is usually a leading indicator of what it builds next.

13

Read next

All 5
14

Open vs closed

Open weights
42%OPEN
36 open weights50 proprietary

Top open model: Qwen3.8-Flash-Next from Alibaba. The gap to the frontier is now measured in months, not generations.

Below: every category, in full. 235 entries across 20 categories.

Language Models

See all 50 →

Every LLM you can call or download today — frontier, mid-tier and open weights. · 50 entries · 8 older versions hidden · ranked by capability and adoption · verified 20 Sept 2026

#NameContextInput ($/Mtok)Output ($/Mtok)LicenceStatus
1GPT-6 Astra
OpenAI
1.05M$10$50Proprietarycurrent
2Claude Fable 5.1
Anthropic
1M$10$50Proprietarycurrent
3Claude Fable 5
Anthropic
1M$10$50Proprietarycurrent
4Claude Opus 5
Anthropic
1M$5$25Proprietarycurrent
5GPT-5.6 Sol
OpenAI
1M+Proprietarycurrent
6Gemini 3.1 Pro
Google
1MProprietarycurrent
7Claude Sonnet 5
Anthropic
1M$3$15Proprietarycurrent
8Grok 4.6
xAI
500K$2$6Proprietarycurrent
9Muse Spark 1.3
Meta
1M$1.25$4.25Proprietarycurrent
10Claude Opus 4.8
Anthropic
1M$5$25Proprietarycurrent

40 more in Language Models, including older versions

Open-Weight Models

See all 36 →

Downloadable weights you can self-host. · 36 entries · 4 older versions hidden · ranked by benchmark performance and adoption · verified 16 Sept 2026

#NameContextInput ($/Mtok)Output ($/Mtok)LicenceStatus
1Qwen3.8-Flash-Next
Alibaba
262K native, ~1M extended$0.15$0.47Open weightscurrent
2DeepSeek V4.1 Flash
DeepSeek
1M$0.3$1.2MITcurrent
3DeepSeek V4-Pro
DeepSeek
MITcurrent
4SAtria Dawn Preview
Shanghai AI Laboratory
256KNot publishedNot publishedMITcurrent
5DeepSeek V4 Flash Vision Exp
DeepSeek
128K$0.44$1.32Open weightsdeprecated
6TInkling
Thinking Machines Lab
1MWeights onlyWeights onlyApache 2.0current
7Kimi K3
Moonshot AI
Open weightscurrent
8GLM-5.3-Flash
Zhipu AI
1M$0.15$0.5MITcurrent
9GLM-5.3
Zhipu AI
1M$1.4$4.4Open weightscurrent
10GLM-5.2
Zhipu AI
Open weightscurrent

26 more in Open-Weight Models, including older versions

OCR & Document AI

See all 13 →

Turning pages into structured text. Now dominated by vision language models rather than the classical OCR engines. · 13 entries · ranked by document parsing benchmarks · verified 2 Sept 2026

#NameOmniDocBenchOpen weightsHandlesLicenceStatus
1GLM-OCR
Zhipu AI
94.6YesTables, formulas, handwritingOpen weightscurrent
2dots.ocr
Xiaohongshu
~93YesLayout, 100+ languagesMITcurrent
3Qwen3-VL
Alibaba
~93YesDocuments, charts, videoApache 2.0current
4DeepSeek-OCR
DeepSeek
~92YesDense text, tablesMITcurrent
5Mistral OCR
Mistral AI
~92NoTables, images, equationsProprietarycurrent
6CCohere Parse 5
Cohere
79.2 on ParseBenchNoTables as HTML, forms, diagrams, bounding boxesProprietarycurrent
7olmOCR
Allen Institute for AI
~91YesTables, markdown structureApache 2.0current
8GOT-OCR 2.0
StepFun
~90YesFormulas, music, chartsApache 2.0current
9MinerU
OpenDataLab
~90YesPDF, formulas, tablesAGPL-3.0current
10Docling
IBM
~88YesPDF, DOCX, PPTX, HTMLMITcurrent

3 more in OCR & Document AI

Object Detection

See all 11 →

Finding and boxing things in images and video. The licence matters as much as the accuracy here. · 11 entries · ranked by COCO mAP, and licence realism · verified 13 Aug 2026

#NameCOCO mAPSpeedLicenceFamilyStatus
1RF-DETR
Roboflow
60.5Real timeApache 2.0Transformer, set predictioncurrent
2YOLO26
Ultralytics
~55Very highAGPL-3.0Single stagecurrent
3YOLOv12
Ultralytics
~55Very highAGPL-3.0Single stagecurrent
4RTMDet
OpenMMLab
~52300+MITSingle stagecurrent
5IGrounding DINO
IDEA Research
~52 zero-shotModerateApache 2.0Open vocabularycurrent
6YOLO-World
Tencent
~46 zero-shotHighGPL-3.0Open vocabularycurrent
7IDINO-X
IDEA Research
~56ModerateProprietaryOpen vocabularycurrent
8SAM 2
Meta
Not applicableReal timeApache 2.0Segmentationcurrent
9OWLv2
Google
~45 zero-shotModerateApache 2.0Open vocabularycurrent
10DETR
Meta
~44LowApache 2.0Transformer, set predictioncurrent

1 more in Object Detection

Image Models

Category page →

Text-to-image and editing. · 8 entries · ranked by Artificial Analysis Arena Elo · verified 13 Aug 2026

#NameArena EloOpen weightsMax resolutionEditingStatus
1GPT Image 2
OpenAI
1370NoYescurrent
2Reve 2.1
Reve
1324Nocurrent
3NNano Banana 2
Google
1322NoBest in classcurrent
4MAI-Image-2.5
Microsoft AI
1308Nocurrent
5Seedream 5.0 Pro
ByteDance
1282Nocurrent
10Qwen Image 2.0 Pro
Alibaba
1232Partialcurrent
12BFLUX.2 max
Black Forest Labs
1229YesYescurrent
MMidjourney V8.2
Midjourney
NoYescurrent

Video Models

Category page →

Text-to-video and image-to-video. · 9 entries · ranked by Artificial Analysis Arena Elo · verified 2 Sept 2026

#NameArena EloNative audioMax lengthOpen weightsStatus
1Gemini Omni Flash
Google
1323Yescurrent
2MiniMax H3
MiniMax
1303Yescurrent
3Seedance 2.0
ByteDance
1264Yescurrent
4LLTX-2.5
Lightricks
Not ratedYes, generated jointly10s per shot, multishot scenesYescurrent
5Kling 3.0 Omni
Kuaishou
Yes, nativecurrent
6Veo 3.1
Google
Yescurrent
7RRunway Gen-4.5
Runway
Partialcurrent
8Wan 2.7
Alibaba
1161YesYescurrent
Sora 2
OpenAI
Yesdeprecated

Voice & Audio

Category page →

TTS, cloning, speech-to-speech. · 6 entries · ranked by Artificial Analysis Speech Arena Elo · verified 13 Aug 2026

#NameArena EloPrice ($/M chars)Clone fromOpen weightsStatus
1Simba 3.2
Speechify
1231$1010sNocurrent
2Qwen-Audio-3.0-TTS-Plus
Alibaba
1230current
3Eleven v3
ElevenLabs
1178$100Nocurrent
4Cartesia Sonic 3
Cartesia
3sNocurrent
5Fish Audio S2 Pro
Fish Audio
1123$15Yescurrent
6Chatterbox
Resemble AI
5sYescurrent

Music Models

Category page →

Song and instrumental generation. · 5 entries · ranked by output quality and API availability · verified 20 Sept 2026

#NameSung vocalsPublic APIExportStatus
1Suno v6
Suno
YesNoYescurrent
2Suno v5
Suno
YesNoYescurrent
3Lyria 3 Pro
Google
YesYesYescurrent
4Eleven Music
ElevenLabs
YesYesYescurrent
5UUdio
Udio
YesNoNocurrent

Talking heads and dubbing. · 3 entries · ranked by quality, licence and real-time capability · verified 13 Aug 2026

#NameLicenceReal-timeMax resolutionStatus
1OMuseTalk
Open source
MITYescurrent
2SSync 1.9
Sync Labs
CommercialNocurrent
3OHallo2
Open source
MITNo4Kcurrent

Agentic Tools

Category page →

Coding agents and work agents. · 7 entries · ranked by agentic coding performance · verified 13 Aug 2026

#NameSurfacePriceSubagentsVendorStatus
1Claude Code
Anthropic
CLI + IDE + webYescurrent
2Codex
OpenAI
CLI + cloudNocurrent
3Cursor
Cursor
IDE$20/mocurrent
4Devin
Cognition
Cloudfrom $20/mocurrent
5GitHub Copilot
GitHub
IDE$10/mocurrent
6Antigravity
Google
IDE + CLIcurrent
7Claude Cowork
Anthropic
Webfrom $20/mocurrent

Agent Frameworks

Category page →

Orchestration libraries and SDKs. · 6 entries · ranked by production maturity and adoption · verified 13 Aug 2026

#NameModelLanguageMCP supportStatusStatus
1LangGraph
LangChain
GraphPython, JSActivecurrent
2SDSPy
Stanford NLP
CompilerPythonActivecurrent
3CrewAI
CrewAI
Role-basedPythonYesActivecurrent
4Claude Agent SDK
Anthropic
HarnessPython, TSYesActivecurrent
5Pydantic AI
Pydantic
TypedPythonActivecurrent
AutoGen
Microsoft
ConversationPythonMaintenancedeprecated

Routers & Gateways

See all 18 →

One API across many models. · 18 entries · ranked by adoption and measured overhead · verified 13 Aug 2026

#NameAdded latencyHostingCachingStatus
1OpenRouter
OpenRouter
40–55msManagedcurrent
2BLiteLLM
BerriAI
Self-hostedSelfRediscurrent
3Requesty
Requesty
8ms P50Managedcurrent
4Portkey
Portkey
8–20msBothSemanticcurrent
5Cloudflare AI Gateway
Cloudflare
EdgeManagedcurrent
6HHelicone
Helicone
LowBothcurrent
7VVercel AI Gateway
Vercel
LowManagedcurrent
8TrueFoundry
TrueFoundry
LowBothcurrent
9Kong AI Gateway
Kong
VariesSelfcurrent
10Not Diamond
Not Diamond
LowManagedcurrent

8 more in Routers & Gateways

Vector Databases

See all 12 →

Where embeddings live, and the layer every retrieval system is built on. · 12 entries · ranked by adoption, then measured latency · verified 13 Aug 2026

#NameIndexHybrid searchHostingLicenceStatus
1pgvector
PostgreSQL
HNSW, IVFFlatYes, with SQLSelf-host or any managed PostgresPostgreSQLcurrent
2Qdrant
Qdrant
HNSWYesSelf-host and managedApache 2.0current
3Pinecone
Pinecone
ProprietaryYesManaged onlyProprietarycurrent
4Milvus
Zilliz
HNSW, IVF, DiskANN, GPUYesSelf-host and managedApache 2.0current
5Weaviate
Weaviate
HNSWYes, nativeSelf-host and managedBSD-3current
6Chroma
Chroma
HNSWLimitedEmbedded, self-host, managedApache 2.0current
7turbopuffer
turbopuffer
Proprietary, object storageYesManaged onlyProprietarycurrent
8LanceDB
LanceDB
IVF-PQ, HNSWYesEmbedded and managedApache 2.0current
9Vespa
Vespa.ai
HNSW plus invertedYes, nativeSelf-host and managedApache 2.0current
10Redis
Redis
HNSW, FLATYesSelf-host and managedRSALv2 / SSPLcurrent

2 more in Vector Databases

RAG & Knowledge Graphs

See all 14 →

Retrieval, augmentation and generation, plus the graph approaches that connect facts rather than merely matching them. · 14 entries · ranked by fitness for retrieval work, then adoption · verified 13 Aug 2026

#NameKindGraphLanguageLicenceStatus
1LlamaIndex
LlamaIndex
FrameworkYes, property graph indexPython, TypeScriptMITcurrent
2LangChain & LangGraph
LangChain
FrameworkVia integrationsPython, TypeScriptMITcurrent
3Haystack
deepset
FrameworkVia integrationsPythonApache 2.0current
4RAGFlow
InfiniFlow
PlatformYesPythonApache 2.0current
5GraphRAG
Microsoft
LibraryYes, corePythonMITcurrent
6HLightRAG
HKU Data Science
LibraryYes, corePythonMITcurrent
7Graphiti
Zep
LibraryYes, temporalPythonApache 2.0current
8Neo4j
Neo4j
DatabaseYes, nativeCypherGPL-3.0 / commercialcurrent
9Cognee
Cognee
FrameworkYes, corePythonApache 2.0current
10UUnstructured
Unstructured
LibraryNoPythonApache 2.0current

4 more in RAG & Knowledge Graphs

Discovery, weights and demos. · 6 entries · ranked by catalogue size and adoption · verified 6 Sept 2026

#NameCatalogueLive demosHosted inferenceStatus
1Hugging Face
Hugging Face
3m+ modelsYes — SpacesYescurrent
2Replicate
Cloudflare
50,000+YesYescurrent
3fal.ai
fal
CuratedYesYescurrent
4ModelScope
Alibaba
LargeYesYescurrent
5RRunware
Runware
400,000+PartialYescurrent
6Civitai
Civitai
Image onlyYesYescurrent

On-demand and serverless compute. · 6 entries · ranked by price-performance and availability · verified 13 Aug 2026

#NameH100 ($/hr)B200 ($/hr)ServerlessSLAStatus
1RunPod
RunPod
$2.34$5.98YesYescurrent
2Lambda
Lambda Labs
$3.99$3.49NoYescurrent
3CoreWeave
CoreWeave
$5.5NoYescurrent
4Modal
Modal
YesYescurrent
5Vast.ai
Vast.ai
NoNocurrent
6Nebius
Nebius
NoYescurrent

GPU & AI Chips

Category page →

Accelerator silicon. · 8 entries · ranked by capability and availability · verified 10 Sept 2026

#NameMemoryBandwidthPriceVendorStatus
1B300 Blackwell Ultra
NVIDIA
288GB HBM3e8 TB/s~$40–50K (verify)NVIDIAcurrent
2B200
NVIDIA
192GB HBM3e$30–50KNVIDIAcurrent
3MI400
AMD
432GB HBM419.6 TB/sAMDcurrent
4HAscend 950DT
Huawei
144GB HiZQ 2.04 TB/sAbove ¥250,000 (about $37K), reportedHuaweicurrent
5H100
NVIDIA
80GB HBM3$25–40KNVIDIAcurrent
6TPU v7 Ironwood
Google
192GB HBM3e7.37 TB/sGooglecurrent
7Trainium 3
AWS
144GB HBM3eAWScurrent
8Groq LPU
Groq
Groqcurrent

Local AI Hardware

Category page →

Machines for running models at home. · 6 entries · ranked by tokens per second per dollar · verified 13 Aug 2026

#NameMemoryBandwidthPrice70B tok/sStatus
1Framework Desktop
Framework
128GB unified~256 GB/s$1,99912–15current
2Mac Studio M3 Ultra
Apple
up to 512GB546–819 GB/s$3–6K20–30current
3RTX 5090
NVIDIA
32GB GDDR71,792 GB/s$2–3Kcurrent
4RTX 3090 (used)
NVIDIA
24GB~$900current
6RTX PRO 6000 Blackwell
NVIDIA
96GB$13,250–20,000current
9DGX Spark
NVIDIA
128GB LPDDR5x273 GB/s$4,6992.7current

Who funds the field. · 6 entries · ranked by AI capital deployed · verified 13 Aug 2026

#NameTypeFund sizeNotableStatus
1AMGX
Abu Dhabi
Sovereign$49BOpenAI, Anthropiccurrent
2Thrive Capital
United States
VentureOpenAI, Databrickscurrent
3Andreessen Horowitz
United States
Venture$90B AUMcurrent
4Lightspeed
United States
VentureAnthropiccurrent
5NVIDIA / NVentures
United States
CorporateOpenAIcurrent
6SoftBank
Japan
CorporateOpenAIcurrent

Who governs the field. · 5 entries · ranked by enforcement scope · verified 13 Aug 2026

#NameJurisdictionInstrumentEnforcing sinceStatus
1EEU AI Office
European Commission
European UnionEU AI Act2 Aug 2026current
2UUK AI Security Institute
United Kingdom
United KingdomSectoral2024current
3CAISI
United States
United StatesNIST AI RMF2024current
4CAC
China
ChinaCybersecurity Law1 Jan 2026current
5MIntl. Network for AI Measurement
Multilateral
MultilateralShared evaluations2024current

Who named the field, who builds it now, who teaches it, and who argues about where it goes, with their books, channels, careers and who taught whom

Sam AltmanDario AmodeiDemis HassabisElon MuskJensen HuangIlya SutskeverMira MuratiGreg BrockmanMustafa SuleymanYann LeCunArthur MenschAlexandr WangEmad MostaqueSundar Pichai+158

Side-by-side spec sheets for the questions people actually ask

Changed this week

Every edit is dated and carries its source

18 SeptQwen3.8-Omni-Flash released, text, image, audio and video into 1M tokens, API onlyNew
18 SeptStep 5 Preview released, 600B sparse MoE at $1 and $2.70 per million tokensNew
15 SeptGemini 3.8 Live released, first on the Artificial Analysis speech to speech index at 82.6New
14 SeptSiri AI ships on iOS 27, built with Google and its Gemini modelsUpdated
14 SeptSAtria Dawn Preview released, 744B agentic weights under the MIT licenceNew
09 SeptSuno v6 released, first Suno model trained on licensed cataloguesNew