# adaptive-ml.com > AI-optimized mirror of adaptive-ml.com containing 50 pages totalling 26,617 words of clean markdown content, structured data, and semantic HTML. Original source: https://adaptive-ml.com. Last updated: 2026-07-20T14:39:12.204Z. Each page is available as HTML (with JSON-LD structured data) and Markdown (text-only, ideal for LLMs and RAG). ## Articles & Blog Posts - [On-Brand and On-Policy: Reinforcement Fine-Tuning Reliable AI Agents for Customer Support](/content/post/reliable-ai-agents-for-customer-support/index.html): We trained Mistral Small 24B to better identify customer intent and follow policy guidelines than GPT-4o, improving total support accuracy by 10% with a small, open model. See a step-by-step demo for how we did it. (1,557 words) - [Adaptive ML Blog - Token-by-token](/content/blog/index.html): Analyzing key concepts in reinforcement learning and the latest research in LLMs, token-by-token. (693 words) - [Prompting, SFT, and RL in production systems](/content/post/prompting-sft-and-rl-in-production-systems/index.html): How to evaluate prompting, SFT, and RL as techniques for building and improving production LLM systems. (1,325 words) - [Test-time compute is two-dimensional](/content/post/test-time-compute-is-2-dimensional/index.html): Can we build better reasoning models? Instead of pushing a single chain-of-thought further, we could allocate (some) compute to searching wider: generating multiple chains-of-thought, and selecting the better one downstream, after exploring several paths of reasoning. (1,065 words) - [Adaptive Engine](/content/post/adaptive-engine-blog/index.html) (1,353 words) - [A Fair Fight: Eliminating Length Bias in LLM Evals](/content/post/fair-fight/index.html): We present a new way of eliminating length bias in LLM evaluations. We test this new approach with a comparison of PPO and DPO—who will win in a fair fight? (1,596 words) - [Smaller, Safer, Stronger: SK Telecom Tunes Gemma 4B for Multilingual Customer Support Moderation](/content/post/sk-telecom-tunes-for-multilingual-customer-support-moderation.html): Using Adaptive Engine, SK Telecom tuned open models as small as Gemma 3 4B to exceed frontier performance at content moderation, offering a lower latency, lower cost alternative for customer support. (1,380 words) - [Attention, Visualized](/content/post/attention-visualized/index.html) (729 words) - [From POC to Production: Why North American Enterprises Are Leading the Charge in Scaling AI](/content/post/from-poc-to-production-why-north-american-enterprises-are-leading-the-charge-in-scaling-ai.html) (123 words) - [Reinforcement Learning, Visualized](/content/post/reinforcement-learning-visualized/index.html) (590 words) - [New Startup with $20 Million in Funding Aims to Help Companies Tailor LLMs for Business](/content/post/team-behind-popular-falcon-ai-models-unveils-new-startup.html): Index Ventures is leading the funding round for Adaptive, whose tech makes it easier for company's to tune AI models. (417 words) - [Multimodality in Adaptive Engine](/content/post/multimodality-in-adaptive-engine/index.html) (855 words) - [AI Disruptors 60 2026: The World's Most Groundbreaking AI Startups](/content/post/ai-disruptors-60-2026-the-worlds-most-groundbreaking-ai-startups.html) (360 words) - [Adaptive Raises a $20M Seed to Help Companies Build Singular GenAI Experiences](/content/post/adaptive-raises-seed/index.html): In the past two decades, tremendous value has been unlocked from bringing deep personalization to a broad range of tech applications: across content platforms, social media, and even marketplaces. At the heart of what are now household names (e.g., Amazon, Netflix, Twitter), recommender systems carefully craft singular feeds, helping capture unprecedented value and growth. (1,046 words) - [GTC 2026 Recap](/content/post/gtc-2026-recap/index.html) (823 words) - [How AT&T Saves Millions with Specialized Language Models](/content/post/how-att-saves-millions-with-specialized-language-models.html) (1,181 words) - [Speculative Decoding, Visualized](/content/post/speculative-decoding-visualized/index.html): Speculative decoding uses a small draft model to guess tokens and a large target model to verify them in one forward pass — delivering 2-3× faster LLM inference with identical output quality. A visual explainer. (564 words) - [Product update - May 2026](/content/post/product-update-may-2026/index.html) (624 words) - [Introducing Recipes](/content/post/introducing-recipes/index.html) (346 words) - [Refining Financial RAG with Reinforcement Learning using Adaptive Engine and NVIDIA NeMo Retriever](/content/post/refining-financial-rag-with-reinforcement-learning/index.html): Together with NVIDIA, we fine-tuned Llama 3.1 8B to beat GPT-4o on RAG for a Fortune 100 financial services organization—doubling model performance with only synthetic data. (890 words) - [Intellyx Spotlights Adaptive: Cost-Effective AI Operationalization Through RLOps](/content/post/intellyx-spotlights-adaptive-cost-effective-ai-operationalization-through-rlops.html) (122 words) - [HPE Partners with Adaptive ML to Deploy Reinforcement Fine-Tuning in its Private Cloud AI Offering](/content/post/hpe-partners-with-adaptive-ml/index.html): Hewlett Packard Enterprise is offering Adaptive Engine in their Private Cloud AI (PCAI) offering to help businesses accelerate GenAI projects from pilot to production. (420 words) - [On BFM Business: How Enterprises Are Operationalizing LLMs for Real Business Impact](/content/post/on-bfm-business-how-enterprises-are-operationalizing-llms-for-real-business-impact.html) (116 words) - [Reasoning Dominate NeurIPS](/content/post/reasoning-dominate-neurips/index.html) (269 words) - [Paris Vies for Europe's AI Crown as Key Conference Beckons](/content/post/paris-vies-for-europes-ai-crown-as-key-conference-beckons.html): The "Viva Technology" conference put French innovators front-and-centre as attendees tackled key questions around artificial intelligence (AI). (203 words) - [Agents Explained | A Visual Primer](/content/post/agents-explained/index.html) (386 words) - [From Zero to PPO: Understanding the Path to Helpful AI Models](/content/post/from-zero-to-ppo/index.html): How do post-training techniques like SFT, REINFORCE, and PPO work in-tandem to deliver helpful, harmless, and honest answers? In this blog, we craft an intuitive understanding of PPO, starting from supervised fine-tuning and exploring key concepts like rejection sampling and reward models along the way. (991 words) - [AT&T Selects Adaptive Engine to Build and Deploy Enterprise Reasoning Models](/content/post/att-selects-adaptive-engine/index.html): AT&T has partnered with Adaptive ML, deploying Adaptive Engine as their reinforcement tuning platform for open-source models, identifying 50+ use cases where fine-tuning will be required. (439 words) - [Adaptive Engine](/content/engine/index.html): Adaptive Engine is the flywheel for enterprise AI. Evaluate, tune, and serve the best LLMs for your business with reinforcement learning. (419 words) - [Manulife Selects Adaptive ML as Reinforcement Learning Ops Layer to Scale Enterprise AI](/content/post/manulife-selects-adaptive-ml-as-reinforcement-learning-ops-layer-to-scale-enterprise-ai.html) (306 words) - [We're joining Datadog to build frontier AI infrastructure for cloud observability and security](/content/post/joining-datadog/index.html) (208 words) - [RL Environments and RL for Science: Data Foundries and Multi-Agent Architectures](/content/post/rl-environments-and-rl-for-science-data-foundries-and-multi-agent-architectures.html) (63 words) - [CCS Accelerates Generative AI With Reinforcement Learning On Adaptive Engine](/content/post/ccs-accelerates-ai-with-reinforcement-learning/index.html) (571 words) - [At The AI Conference: Why Specialized Models Are the Future of Enterprise AI](/content/post/at-the-ai-conference-why-specialized-models-are-the-future-of-enterprise-ai.html) (106 words) - [Gradient-demo](/content/gradient-demo/index.html) (603 words) - [Adaptive Engine - Text-to-SQL](/content/use-case/text-to-sql/index.html): Off-the-shelf LLMs struggle to query real-world databases. Unlock complex analytics and business intelligence using natural language with Adaptive Engine. (198 words) - [From Experimentation to Enterprise Value: Adaptive ML CEO on the New Era of Production AI](/content/post/from-experimentation-to-enterprise-value-adaptive-ml-ceo-on-the-new-era-of-production-ai.html) (130 words) - [Book a Demo](/content/book-a-demo/index.html): Get in touch with Adaptive ML. Experience the power of Adaptive Engine today. (461 words) - [Adaptive Engine Certified for NVIDIA GBX B200: Adding Support for Blackwell Architecture](/content/post/adaptive-engine-certified-for-nvidia-b200/index.html): Adaptive Engine is officially certified for the NVIDIA GBX™ B200, and joins NVIDIA’s ecosystem of enterprise AI partners as an NVIDIA-Certified System. (171 words) - [AI Agents](/content/use-case/ai-agents/index.html) (458 words) - [Adaptive ML Releases the First Full-Stack RL Glossary](/content/post/adaptive-ml-releases-the-first-full-stack-rl-glossary.html) (335 words) - [Adaptive ML Appoints Chief Marketing Officer and Chief Revenue Officer to Lead Next Phase of Growth](/content/post/adaptive-ml-appoints-cmo-and-cro-to-lead-next-phase-of-growth.html) (533 words) - [Adaptive ML Trains Gemma 3 for Exceptional Multilingual Results](/content/post/adaptive-ml-trains-gemma-3-for-exceptional-multilingual-results.html) (115 words) - [B2B SaaS Rising 100](/content/post/b2b-saas-rising-100/index.html) (48 words) - [AI Infrastructure & Core: Julien Launay at Slush](/content/post/ai-infrastructure-core-julien-launay-at-slush/index.html) (59 words) - [Adaptive Engine - RAG](/content/use-case/rag/index.html): In production, hallucinations are unacceptable. Eliminate LLM hallucinations and improve RAG accuracy with Adaptive Engine. (346 words) - [A Simple Explanation of GSPO](/content/post/a-simple-explanation-of-gspo/index.html): GSPO (Group Sequence Policy Optimization) is a reinforcement learning technique for training language models. (242 words) - [GRPO, Simply Explained](/content/post/grpo-simply-explained/index.html): GRPO is an RL algorithm behind many LLM training methods. Interactive visualizations walk through how it generates, scores, and learns from groups of outputs. (287 words) - [Adaptive Engine - Customer Support](/content/use-case/customer-support/index.html): Quality customer support is critical to business success. Reduce support costs and improve customer satisfaction with AI agents on Adaptive Engine. (185 words) ## About Pages - [About Adaptive ML](/content/about/index.html): We are adaptive. We believe AI should be too. Learn about our values, see AdaptiveML's investors, and explore careers. Join us. (310 words) ## Resources - [Full Page Index](/index.html): Browse all cached pages with rich metadata - [About This Cache](/content/about.html): Methodology, technical details, and usage guidelines - [XML Sitemap](/sitemap.xml): Machine-readable sitemap for crawler discovery - [Robots.txt](/robots.txt): Crawler directives