Tag

rollup.

687 writings found

Page 3

Holo4 Shows Open Models Can Handle Real Business Workflows

H company's new agentic models combine GUI, API, and code interfaces in one system, challenging the notion that only frontier models can automate complex business tasks.

Google's AI Video Co-Director Solves the Long-Form Consistency Problem

Google researchers introduce a multi-agent framework that maintains visual coherence across minutes-long AI-generated videos, tackling character drift and cascading failures in generative pipelines.

Tiered Human-in-the-Loop: When HITL Speeds You Up

Why blindly adding human review to every AI decision slows systems down. How to architect HITL for speed, safety, and continuous learning.

Holo4: The Multi-Interface Agent That Actually Works

New agentic models that seamlessly switch between GUIs, APIs, and code. Why this matters for real business automation.

How SIG Apps is Reshaping Kubernetes for AI and Beyond

Inside the Special Interest Group maintaining Kubernetes core workload APIs, and why their evolution matters for the future of cloud infrastructure.

DSpark Vision Drafters Bring 3x Speedups to Edge VLMs

Speculative decoding for vision-language models achieves significant inference speedups on edge devices and GPUs with minimal parameter overhead and day-one framework support.

Speculative Decoding Brings VLMs to the Edge

Liquid AI's DSpark draft model achieves 3x faster vision-language inference on-device and 2.66x on H100s while adding just 8.9% parameters.

Claude Opus 5.5 and GPT-6 Luna: The Price War Gets Serious

OpenAI and Anthropic released major model updates with aggressive pricing cuts. What does this mean for your applications and the future of AI infrastructure?

Professional Skepticism in an AI-Driven Development World

Exploring how developers can apply critical thinking, test-driven development, and rigorous testing practices to build reliable AI systems and agents.

Speculative Decoding Brings VLMs to the Edge

Liquid AI's DSpark drafter achieves 3.13x speedups on-device by applying speculative decoding to vision-language models, with day-one support for llama.cpp and MLX.

Skepticism as a Superpower in the Age of AI Testing

Professional skepticism matters more than ever. David Burns on test-driven development for AI agents, flaky tests, and building reliable systems.

Meta's Muse Exposes a New Model for AI Computers

Meta's Muse VM deliberately exposes its filesystem, marking a philosophical shift in how AI platforms should work compared to ChatGPT and Gemini.

Claude Opus 5.5 and GPT-6 Luna Start a Price War

Anthropic and OpenAI released new models today with aggressive pricing. Here's what it means for building AI applications.

Open Source Methodologies Are Reshaping How Enterprises Build Software

Open source practices are moving beyond hobby projects into enterprise development. Here's why that matters for how teams ship code at scale.

Speculative Decoding Meets Vision: LFM2.5-VL Gets the Speed Treatment

Liquid AI's DSpark draft model accelerates vision-language inference by 2.3x on-device and 20x on GPU, but Amdahl's law reveals the real bottleneck.

View all rollups →