Category
Technologies
LLM Articles
Keep up to date with the latest techniques, tools, and research in Large Language Models. Our blog talks about data science, uses, & responsible AI practices.
Other technologies:
Training 2 or more people?Try DataCamp for Business
Best LLM for Coding in 2026: 9 Models Ranked
We rank the 9 best coding LLMs of September 2026, from Claude Opus 5.5 to open-weight models you can run locally, using SWE-bench Pro and Terminal-Bench.
Tim Lu
September 25, 2026
Top 7 Open-Source TypeSafe Jev Alternatives
Explore a new wave of open-source projects offering practical Jev alternatives for fast, local, and cost-efficient AI decision-making.
Abid Ali Awan
September 23, 2026
GPT-6 Sol vs Claude Opus 5.5: Benchmarks, Pricing, and Which to Use
Two same-day efficiency releases, priced two to one. We compare GPT-6 Sol and Claude Opus 5.5 on benchmarks, real workload costs, access, and a hands-on Tetris build.
Tom Farnschläder
September 23, 2026
The History and Evolution of Large Language Models: From ELIZA to AI Agents
In 1966, the chatbot ELIZA could do little more than rephrase a user’s statements as questions. Sixty years later, large language models can plan, write code, and use tools on their own. Here, we will walk through the history of LLMs across 8 eras, and what each breakthrough actually changed.
Vaibhav Mehra
September 22, 2026
Grok 4.7: Features, Benchmarks, Pricing, and Access
xAI's Grok, a new coding and knowledge-work model, runs at $2/$6 per million tokens and tops frontier rivals on some benchmarks. Here's what's new, how it compares to the competition, and how it upgrades on its predecessor.
Matt Crabtree
September 21, 2026
TypeSafe Jev vs GPT-6 Astra: Decision Model or Frontier Agent?
Jev decides in milliseconds and cannot write a word; GPT-6 Astra finishes whole tasks unattended. Here is which pipeline steps belong to each model, and what they cost.
Tom Farnschläder
September 21, 2026
Bonsai 2 27B: Run a 27B AI Locally on Your Laptop
PrismML's Bonsai 2 27B compresses a 27B multimodal model to 5.9GB while keeping 98.2% of full-precision benchmark performance, and it runs on a laptop.
Matt Crabtree
September 18, 2026
Qwen3.8-Omni-Flash: Alibaba's Omnimodal Agent Model
Alibaba's Qwen3.8-Omni-Flash handles text, image, audio, and video with a 1M-token context window, beating Gemini 3.8 Flash on audio benchmarks.
Matt Crabtree
September 18, 2026
DeepSeek V4.1 Flash vs Gemini 3.8 Flash: Benchmarks, Pricing, and Which to Use
DeepSeek's open-weight V4.1 Flash matches Gemini 3.8 Flash on agentic coding benchmarks at a third of the price. We compare specs, costs, and a hands-on build test.
Tom Farnschläder
September 17, 2026
Jev: TypeSafe's System One Model That Never Hallucinates
TypeSafe's Jev is a new class of AI model, a System One Model, that returns typed decisions with calibrated probabilities instead of text, running 40-200x faster than frontier LLMs.
Matt Crabtree
September 16, 2026
ChatGPT Images 2.5: Features, Sketch, Templates, and API Models
OpenAI's new image model brings sharper details, up to 50% faster generation, and new ChatGPT tools like Sketch and templates, plus two API models.
Matt Crabtree
September 9, 2026
5 GPT-6 Astra Projects to Test Out OpenAI’s New Model
Discover five GPT-6 Astra projects that can help you explore the new features and upgrades of OpenAI’s new model.
Matt Crabtree
September 9, 2026