Modellen & prijzen nieuws
Al het nieuws over Modellen & prijzen uit onze dagelijkse nieuwsbrief: 70 berichten, nieuwste eerst. Koppen komen van de bron; de Nederlandse duiding is machinaal geschreven.
zondag 27 september 2026
- 💰 LLM pricing isn’t just about cost per request. It’s about cost per token × input/output tokens × request volume. A model that looks cheape · X · 2026-09-26 · nieuwsbrief
- Tuesday morning AI FACTS roundup 🤖 22.09 - 🚀 xAI unveils the Grok 4.7 model with longer agentic operation and pricing of $2 / $6 per 1M toke · X · 2026-09-22 · nieuwsbrief
- If you use a full max plan at API pricing that’s ~$8k/mo or $96k/year. Thats why most big companies who can’t get max plans are slashing LLM · X · 2026-09-26 · nieuwsbrief
- refactor(llm): replace litellm model-cost table with vendored catalog · GitHub · 2026-09-22 · nieuwsbrief
zaterdag 26 september 2026
- Tuesday morning AI FACTS roundup 🤖 22.09 - 🚀 xAI unveils the Grok 4.7 model with longer agentic operation and pricing of $2 / $6 per 1M toke · X · 2026-09-22
xAI introduceert het Grok 4.7‑model met langere agent‑operatie en een prijs van $2/$6 per miljoen tokens, relevant voor indie‑developers die betaalbare LLM‑capaciteit zoeken. · nieuwsbrief
vrijdag 25 september 2026
- refactor(llm): replace litellm model-cost table with vendored catalog · GitHub · 2026-09-22 · nieuwsbrief
- Tuesday morning AI FACTS roundup 🤖 22.09 - 🚀 xAI unveils the Grok 4.7 model with longer agentic operation and pricing of $2 / $6 per 1M toke · X · 2026-09-22 · nieuwsbrief
- Global LLM price war continues - OpenAI has introduced two new tiers—GPT‑6 Sol and GPT‑6 Luna—to its GPT‑6 series, with API pricing reduced · X · 2026-09-23 · nieuwsbrief
- LiteLLM replacement G3: gateway streaming + usage/spend recording (UNKNOWN never 0) · GitHub · 2026-09-24 · nieuwsbrief
dinsdag 22 september 2026
- docs: add a model loading and startup acceleration guide · GitHub · 2026-09-21
docs: add a model loading and startup acceleration guide voegt een gids toe voor sneller model‑laden, wat indie‑ontwikkelaars tijd kan besparen bij het opstarten. · nieuwsbrief - [openrouter] Missing model: google/gemini-3.8-flash · GitHub · 2026-09-17
[openrouter] Missing model: google/gemini-3.8-flash meldt dat het model google/gemini‑3.8‑flash niet beschikbaar is, wat indie‑ontwikkelaars moet laten weten dat ze een alternatief moeten zoeken. · nieuwsbrief - How to Write with an LLM · Hacker News · 2026-09-17
How to Write with an LLM biedt richtlijnen voor het schrijven met een groot taalmodel, wat indie‑ontwikkelaars kan ondersteunen bij content‑generatie. · nieuwsbrief - [FEATURE] configure model deployments with non standard names to use the standard name for cost and capabilities discovery · GitHub · 2026-09-17
[FEATURE] configure model deployments with non standard names to use the standard name for cost and capabilities discovery beschrijft een functie om niet‑standaard modelnamen te koppelen aan standaardnamen voor kosten‑ en capaciteitsdetecti · nieuwsbrief
zondag 20 september 2026
- Deploy an LLM as an agent instead of a prompt, and watch your token count jump as much as 10-100x. Nothing about the model changes. Nothing · X · 2026-09-17
Het artikel waarschuwt dat het inzetten van een LLM als agent i.p.v. prompt het token‑verbruik sterk kan verhogen, wat belangrijk is voor indie‑developers met een beperkt budget. · nieuwsbrief - [openrouter] Missing model: google/gemini-3.8-flash · GitHub · 2026-09-17
De melding over het ontbrekende model google/gemini‑3.8‑flash laat indie‑developers weten dat ze een alternatief moeten zoeken of de fout moeten afhandelen. · nieuwsbrief - How to Write with an LLM · Hacker News · 2026-09-17
De gids legt uit hoe je met een LLM kunt schrijven, wat indie‑developers kan ondersteunen bij het genereren van tekst voor hun projecten. · nieuwsbrief - The turbo-fast jump in LLM quality we’ve seen over the past year doesn’t actually move the AI industry the way most people think. It’s not l · X · 2026-09-18
Het artikel stelt dat de snelle kwaliteitsstijging van LLM’s de industrie niet op de verwachte manier verandert, wat indie‑developers helpt realistische doelen te stellen. · nieuwsbrief
zaterdag 19 september 2026
- The turbo-fast jump in LLM quality we’ve seen over the past year doesn’t actually move the AI industry the way most people think. It’s not l · X · 2026-09-18
De snelle verbetering in LLM‑kwaliteit verandert de AI‑markt minder dan verwacht, wat indie‑developers helpt realistische verwachtingen te stellen. · nieuwsbrief - feat(llm-catalog): add InclusionAI Ling 3.0 Flash VL · GitHub · 2026-09-17
InclusionAI Ling 3.0 Flash VL is toegevoegd aan llm‑catalog, waardoor indie‑developers een nieuw model kunnen gebruiken. · nieuwsbrief - How to Write with an LLM · Hacker News · 2026-09-17
Artikel over hoe je met een LLM kunt schrijven, biedt indie‑developers praktische tips voor contentgeneratie. · nieuwsbrief - input costs, output is free - the opposite of typical llm pricing. · X · 2026-09-18
Deze LLM rekent alleen voor invoer en biedt gratis output, wat indie‑developers kosten kan verlagen. · nieuwsbrief
vrijdag 18 september 2026
- feat(backend): add Unbiased Pareto to LLM catalog · GitHub · 2026-09-18 · nieuwsbrief
- Deploy an LLM as an agent instead of a prompt, and watch your token count jump as much as 10-100x. Nothing about the model changes. Nothing · X · 2026-09-17 · nieuwsbrief
- feat(llm-catalog): add InclusionAI Ling 3.0 Flash VL · GitHub · 2026-09-17 · nieuwsbrief
- How to Write with an LLM · Hacker News · 2026-09-17 · nieuwsbrief
woensdag 16 september 2026
- Building your first LLM API call in Python (step by step) · Hacker News · 2026-09-15 · nieuwsbrief
- ok everyone, i took one for the team here's the data we all wanted to see - real token value of each LLM subscription, empirically measured · X · 2026-09-11 · nieuwsbrief
- Show HN: Determinstic LLM inference for lowest price Gemma 4, with Windows XP · Hacker News · 2026-09-12 · nieuwsbrief
- Admin: model configuration per stage, and what each stage has spent — needs a design session · GitHub · 2026-09-13 · nieuwsbrief
dinsdag 15 september 2026
- ok everyone, i took one for the team here's the data we all wanted to see - real token value of each LLM subscription, empirically measured · X · 2026-09-11 · nieuwsbrief
- Introducing Colossus and the launch of $COLOSSUS 🗽 Colossus brings price discovery to AI inference through a marketplace for discounted LLM · X · 2026-09-08 · nieuwsbrief
- Show HN: Determinstic LLM inference for lowest price Gemma 4, with Windows XP · Hacker News · 2026-09-12 · nieuwsbrief
- Admin: model configuration per stage, and what each stage has spent — needs a design session · GitHub · 2026-09-13 · nieuwsbrief
zondag 13 september 2026
- Real-SWE: Benchmarking AI models on private, real-world, enterprise codebases · Hacker News · 2026-09-12
Een benchmark (Real‑SWE) van AI‑modellen op private, enterprise codebases, relevant voor indie‑developers die modelprestaties willen vergelijken. · nieuwsbrief - Introducing Colossus and the launch of $COLOSSUS 🗽 Colossus brings price discovery to AI inference through a marketplace for discounted LLM · X · 2026-09-08
Introductie van Colossus, een marktplaats voor goedkope LLM‑inference, relevant voor indie‑developers die kosten willen besparen op AI‑gebruik. · nieuwsbrief - Another angle of my LLM coding leaderboard: price/quality. Two charts: 1. Cheaper LLMs up to $0.50 per prompt (API pricing) 2. More expensi · X · 2026-09-10
Een analyse van prijs‑kwaliteit van LLM’s, met charts voor goedkope en duurdere modellen, nuttig voor indie‑developers bij het kiezen van een model. · nieuwsbrief - Admin: model configuration per stage, and what each stage has spent — needs a design session · GitHub · 2026-09-13
Een GitHub‑issue over modelconfiguratie per fase en kostenoverzicht, relevant voor indie‑developers die hun AI‑pipeline willen beheren. · nieuwsbrief
zaterdag 12 september 2026
- 🚀 OpenBMB releases MiniCPM5-2B! A 2.5B parameter dense LLM setting a new 2B-class SOTA (53.9 avg). Features: • 131k context window • Outper · X · 2026-09-08
OpenBMB heeft MiniCPM5-2B uitgebracht, een 2,5 B‑parameter LLM met een contextvenster van 131 k, wat indie‑ontwikkelaars lange‑teksttoepassingen mogelijk maakt. · nieuwsbrief
woensdag 9 september 2026
- Introducing Colossus and the launch of $COLOSSUS 🗽 Colossus brings price discovery to AI inference through a marketplace for discounted LLM · X · 2026-09-08 · nieuwsbrief
- LLM inference pricing changes daily across platforms. Price Dashboard tracks 100 plus providers every 6 hours: - History for DeepSeek, K3 an · X · 2026-09-08 · nieuwsbrief
- → https://StandardCOmpute.com is a flat-rate, OpenAI-compatible LLM API for coding agents designed to lower AI costs via smart model routing · X · 2026-09-08 · nieuwsbrief
- → @FluxionAi_HQ is a unified AI API gateway that provides pay-as-you-go access to global and Chinese LLM models with discounted token pricin · X · 2026-09-08 · nieuwsbrief
zondag 6 september 2026
- We benchmarked Claude Fable 5.1 and Gemini 3.8 Flash on surgical intelligence: - The performance of the two new releases is spiky (e.g. goo · X · 2026-09-02
Het is een benchmark van Claude Fable 5.1 en Gemini 3.8 Flash op chirurgische intelligentie, waarbij de wisselvallige prestaties worden gemeld, wat relevant is voor indie‑ontwikkelaars die deze modellen voor kritieke toepassingen willen geb · nieuwsbrief
vrijdag 4 september 2026
- We benchmarked Claude Fable 5.1 and Gemini 3.8 Flash on surgical intelligence: - The performance of the two new releases is spiky (e.g. goo · X · 2026-09-02
De benchmark vergelijkt Claude Fable 5.1 en Gemini 3.8 Flash op chirurgische intelligentie en meldt wisselvallige prestaties, wat indie‑developers moet overwegen bij modelkeuze. · nieuwsbrief
donderdag 3 september 2026
- We benchmarked Claude Fable 5.1 and Gemini 3.8 Flash on surgical intelligence: - The performance of the two new releases is spiky (e.g. goo · X · 2026-09-02
Een X‑bericht met benchmarkresultaten van Claude Fable 5.1 en Gemini 3.8 Flash op chirurgische intelligentie, relevant voor indie‑developers die modelprestaties vergelijken. · nieuwsbrief
maandag 31 augustus 2026
- Confession: with my LLM coding benchmarks, I'm pretty close to scenario when Terminal is ALWAYS running some evaluation of some model. Almo · X · 2026-08-29
De auteur geeft aan dat hun terminal continu LLM‑evaluaties uitvoert, wat indie‑developers inzicht geeft in runtime‑prestaties. · nieuwsbrief
zondag 30 augustus 2026
- Confession: with my LLM coding benchmarks, I'm pretty close to scenario when Terminal is ALWAYS running some evaluation of some model. Almos · X · 2026-08-29
De tweet geeft aan dat de terminal continu LLM‑evaluaties uitvoert, wat indie‑developers bewust maakt van de mogelijke compute‑belasting. · nieuwsbrief
vrijdag 28 augustus 2026
- -> AI labs spend zillions of dollars on data and compute -> Models get a little better -> "We beat the benchmark" -> Repeat over and over ag · X · 2026-08-21
Korte analyse van de cyclus waarin AI‑labs veel geld uitgeven, modellen verbeteren en benchmarks halen, relevant voor indie‑developers die AI‑kosten overwegen. · nieuwsbrief
woensdag 26 augustus 2026
- feat(aio): use Vercel AI Gateway reported cost for LLM generations · GitHub · 2026-08-25
De aio‑feature maakt gebruik van de door Vercel AI Gateway gerapporteerde kosten voor LLM‑generaties, wat indie‑ontwikkelaars helpt hun uitgaven te monitoren. · nieuwsbrief - If I had 6 months to master Context Engineering. I'd do this. Prompt engineering is dead. Context engineering is the new meta. Stage 1: LLM · X · 2026-08-21
De auteur stelt dat prompt‑engineering achterhaald is en dat context‑engineering de nieuwe focus is, wat indie‑ontwikkelaars kan leiden tot betere LLM‑integraties. · nieuwsbrief - The real value of @OpenRouter is aggregating latency, uptime, throughput, pricing info across the 100+ hosted oss llm providers But you do n · X · 2026-08-25
OpenRouter biedt aggregatie van latency, uptime, doorvoersnelheid en prijsinformatie van gehoste open‑source LLM‑providers — volgens de X-post van meer dan 100 aanbieders, al noemen publieke bronnen circa 400+ modellen van 60+ providers — wat nuttig is voor indie‑ontwikkelaars die providers willen vergelijken. · nieuwsbrief - consumer llm pricing · X · 2026-08-24
Er is informatie over consumenten‑LLM‑prijzen, wat indie‑ontwikkelaars helpt bij kostenplanning. · nieuwsbrief
dinsdag 25 augustus 2026
- Everyone is trying to searching how to save on LLM API costs, but people rarely understand how those API costs are setup in the first place. · X · 2026-08-23
Dit wijst op een gebrek aan begrip over hoe LLM-API-kosten zijn opgebouwd, wat relevant is voor indie-ontwikkelaars die AI-kosten willen beheersen. · nieuwsbrief - OCR It – pull text out of un-copyable documents for your LLM · Hacker News · 2026-08-24
Dit is een OCR-hulpmiddel om tekst uit niet-kopieerbare documenten te halen voor een LLM, wat indie-ontwikkelaars helpt om documentdata aan AI te voeden. · nieuwsbrief - consumer llm pricing · X · 2026-08-24
Dit betreft consumentenprijzen van LLM's, wat indie-ontwikkelaars helpt om de kosten voor eindgebruikers van AI-functionaliteit in te schatten. · nieuwsbrief - DeepSeek V4 Pro is GA: adaptive reasoning, Responses API, peak/off-peak pricing. Mixed benchmarks, but strong in cybersecurity. So what? A s · X · 2026-08-20
Dit kondigt DeepSeek V4 Pro aan met adaptieve redenering, een Responses API, piek- en dalprijzen en gemengde benchmarks, wat indie-ontwikkelaars een extra LLM-optie geeft. · nieuwsbrief
maandag 24 augustus 2026
- -> AI labs spend zillions of dollars on data and compute -> Models get a little better -> "We beat the benchmark" -> Repeat over and over ag · X · 2026-08-21
De beschrijving van de cyclus waarin AI‑labs veel geld uitgeven aan data en compute, laat indie‑developers de context van modelverbeteringen zien. · nieuwsbrief
zondag 23 augustus 2026
- I built a real-time LLM API pricing comparator — because I was tired of not knowing the actual cost difference between models · Reddit · 2026-08-21 · nieuwsbrief
- OpenAI dropped GPT-5.6 Sol pricing on August 20. Input is now $4 per million tokens, output $20 per million, 20% lower input and 33% lower o · X · 2026-08-22 · nieuwsbrief
- Alibaba releases Qwen3.8-27B open weights — a 27B multimodal model that beats Qwen3.7-Plus on several coding and agent benchmarks · Reddit · 2026-08-17 · nieuwsbrief
zaterdag 22 augustus 2026
- OpenAI cuts developer pricing for frontier GPT-5.6 Sol model by more than 20% · Hacker News · 2026-08-22 · nieuwsbrief
- Alibaba releases Qwen3.8-27B open weights — a 27B multimodal model that beats Qwen3.7-Plus on several coding and agent benchmarks · Reddit · 2026-08-17 · nieuwsbrief
- What locked-in price and usage limits would make you switch LLM providers? · Reddit · 2026-08-18 · nieuwsbrief
vrijdag 21 augustus 2026
- We burned 11.7B tokens to find the best cyber AI model · Hacker News · 2026-08-21 · nieuwsbrief
- feat(cognition): give Cognition its own provider identity · GitHub · 2026-08-21 · nieuwsbrief
- feat(travelclick): add travelclick · GitHub · 2026-08-21 · nieuwsbrief
- WhatsApp raised some API rates 300% and will end free messages soon · Hacker News · 2026-08-21 · nieuwsbrief
donderdag 20 augustus 2026
- Most AI cost overruns are not a pricing problem. They are a routing problem. A classification task that could run on a small model runs on a · X · 2026-08-19 · nieuwsbrief
- New benchmark just dropped. Not another "can it solve the end of the world" exam. Not another biology or philosophy olympiad. This one answe · X · 2026-08-19 · nieuwsbrief
- feat(llm): add moonshot/kimi-k3 to model prices and context window map · GitHub · 2026-08-19 · nieuwsbrief