<?xml version="1.0" encoding="UTF-8"?><rss version="2.0"><channel><title>Saurabh Sarkar</title><description>Notes on applied AI, model evaluation, and production systems.</description><link>https://saurabhsarkar.com/</link><item><title>Jev Beat the 4B Model. Then the 4B Model Beat Jev.</title><link>https://saurabhsarkar.com/writing/jev-beat-the-4b-model-then-the-4b-model-beat-jev/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/jev-beat-the-4b-model-then-the-4b-model-beat-jev/</guid><description>Jev beat untouched small LLMs. After split-safe fine-tuning, a 4B model learned the task and transferred better on BeaverTails and C-SafeQA.</description><pubDate>Mon, 28 Sep 2026 00:00:00 GMT</pubDate></item><item><title>Your fraud model has never seen the attack that will work</title><link>https://saurabhsarkar.com/writing/fraud-model-never-seen-the-attack/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/fraud-model-never-seen-the-attack/</guid><description>A fraud model trains on fraud that already worked. When a new ring hit an unsecured lender, the behavioural anomaly model with no fraud labels caught 85% of it and the supervised model caught a third.</description><pubDate>Tue, 22 Sep 2026 00:00:00 GMT</pubDate></item><item><title>Why construction estimates lose their margin after the bid is won</title><link>https://saurabhsarkar.com/writing/why-construction-estimates-lose-margin-after-the-bid/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/why-construction-estimates-lose-margin-after-the-bid/</guid><description>Construction estimates lose margin between bid and closeout because competitive bidding selects the low side of your own estimating error. Here is the mechanism, and the four questions that remeasure the rule.</description><pubDate>Tue, 22 Sep 2026 00:00:00 GMT</pubDate></item><item><title>Why Couldn&apos;t an LLM Pick Its Best Expert?</title><link>https://saurabhsarkar.com/writing/why-couldnt-an-llm-pick-its-best-expert/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/why-couldnt-an-llm-pick-its-best-expert/</guid><description>A hindsight oracle could name the best expert in an LLM. A trained judge, a linear policy and a bounded nonlinear check could not find it live. Why a real opportunity never became a usable one.</description><pubDate>Tue, 22 Sep 2026 00:00:00 GMT</pubDate></item><item><title>Why thin-file approvals look fine in backtest and fail in production</title><link>https://saurabhsarkar.com/writing/why-thin-file-approvals-fail-in-production/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/why-thin-file-approvals-fail-in-production/</guid><description>Thin-file approvals look fine in backtest because repayment labels only exist for applicants you already approved. Here is why the cutoff fails in production, and the four questions that remeasure it.</description><pubDate>Tue, 22 Sep 2026 00:00:00 GMT</pubDate></item><item><title>What travels around the loop?</title><link>https://saurabhsarkar.com/writing/what-travels-around-the-loop/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/what-travels-around-the-loop/</guid><description>We swapped hidden states in Coconut and a small recurrent transformer to test what their loops carry and how they use it.</description><pubDate>Tue, 15 Sep 2026 00:00:00 GMT</pubDate></item><item><title>Deleting a Model&apos;s Thoughts Changed Nothing. Then I Changed the Task.</title><link>https://saurabhsarkar.com/writing/deleting-a-model-s-thoughts-changed-nothing-then-i-changed-the-task/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/deleting-a-model-s-thoughts-changed-nothing-then-i-changed-the-task/</guid><description>Deleting a model&apos;s thoughts barely changed graph-search accuracy but cut arithmetic from 33.3% to 5.1% in this Coconut replication.</description><pubDate>Mon, 14 Sep 2026 00:00:00 GMT</pubDate></item><item><title>I Sent a Small AI Model to School. Here’s Its Report Card.</title><link>https://saurabhsarkar.com/writing/i-sent-a-small-ai-model-to-school-here-s-its-report-card/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/i-sent-a-small-ai-model-to-school-here-s-its-report-card/</guid><description>A Small AI Model spent a school year with a larger teacher. It scored 25/78 on unseen puzzles, versus 8/78 for its unchanged copy. Here’s what stuck.</description><pubDate>Sun, 13 Sep 2026 00:00:00 GMT</pubDate></item><item><title>I forced a Mixture of Experts to specialize by topic. It got worse.</title><link>https://saurabhsarkar.com/writing/forced-mixture-of-experts-specialize-by-topic/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/forced-mixture-of-experts-specialize-by-topic/</guid><description>Five Mixture of Experts models trained from scratch. Routing correlates more with a token&apos;s surface form than with the document&apos;s topic, and the pattern is fixed in the first 2% of training. Forcing topic specialization three ways made every model slightly worse.</description><pubDate>Tue, 28 Jul 2026 00:00:00 GMT</pubDate></item><item><title>What does a Mixture-of-Experts router actually read?</title><link>https://saurabhsarkar.com/writing/what-does-a-mixture-of-experts-router-actually-read/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/what-does-a-mixture-of-experts-router-actually-read/</guid><description>A Mixture of Experts router picks its experts by reading the residual stream. A Jacobian lens on OLMoE and Qwen-MoE shows it sorts token form, brackets and verb tense, not meaning.</description><pubDate>Wed, 22 Jul 2026 00:00:00 GMT</pubDate></item><item><title>Mixture of Experts: the sparse model that lights up 93% of itself</title><link>https://saurabhsarkar.com/writing/mixture-of-experts-the-sparse-model-that-lights-up-93-percent-of-itself/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/mixture-of-experts-the-sparse-model-that-lights-up-93-percent-of-itself/</guid><description>A Mixture of Experts should use a fraction of itself per token. Measured on a 4090: 628 of 1,024 expert slots fire every step, and 93% of them at batch 32.</description><pubDate>Mon, 20 Jul 2026 00:00:00 GMT</pubDate></item><item><title>J-Scope: before a model answers, it thinks in concepts</title><link>https://saurabhsarkar.com/writing/j-scope-before-a-model-answers-it-thinks-in-concepts/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/j-scope-before-a-model-answers-it-thinks-in-concepts/</guid><description>The prompt never says Italy, but twelve layers into Qwen3.5-4B the model is already thinking it. J-Scope lets you watch that happen, inject France and get Paris, or drag one vector and watch the answer walk from euro to yen. A 72-second tour here; the live tool is one click away.</description><pubDate>Wed, 15 Jul 2026 00:00:00 GMT</pubDate></item><item><title>I Fed a Model Facts Through Its J-Space Instead of Its Prompt</title><link>https://saurabhsarkar.com/writing/i-fed-a-model-facts-through-its-j-space-instead-of-its-prompt/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/i-fed-a-model-facts-through-its-j-space-instead-of-its-prompt/</guid><description>I wrote facts directly into a 4B model&apos;s J-space and KV cache on an RTX 4090. Three cache tokens matched full text-RAG, survived 48 distractors, and showed where retrieved facts actually live.</description><pubDate>Wed, 08 Jul 2026 00:00:00 GMT</pubDate></item><item><title>Can you tell a good AI agent plan from a bad one before you run it? I spent 6 million tokens finding out.</title><link>https://saurabhsarkar.com/writing/can-you-tell-a-good-ai-agent-plan-from-a-bad-one-before-you-run-it-i-spent-6-million-tokens-finding/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/can-you-tell-a-good-ai-agent-plan-from-a-bad-one-before-you-run-it-i-spent-6-million-tokens-finding/</guid><description>Can a cheap reward model tell a good AI agent plan from a bad one before you run it? I reproduced Orch-RM on a consumer RTX 4090 and RTX 3090: the verifier barely beats a coin flip on the signal that matters, yet still beats majority vote. Cheap to run, expensive to teach.</description><pubDate>Thu, 02 Jul 2026 00:00:00 GMT</pubDate></item><item><title>I Gave My Coding Agent Karpathy&apos;s Discipline Rules. It Got Too Careful to Fix the Bug.</title><link>https://saurabhsarkar.com/writing/i-gave-my-coding-agent-karpathy-s-discipline-rules-it-got-too-careful-to-fix-the-bug/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/i-gave-my-coding-agent-karpathy-s-discipline-rules-it-got-too-careful-to-fix-the-bug/</guid><description>I tested the viral Karpathy coding-discipline rules on 100 real SWE-bench bugs, three times. They didn&apos;t help: the agent fixed fewer of them, trading fixes for caution.</description><pubDate>Sun, 28 Jun 2026 00:00:00 GMT</pubDate></item><item><title>I bolted MiniMax&apos;s MSA sparse attention onto a 3B model on a single 4090</title><link>https://saurabhsarkar.com/writing/i-bolted-minimax-s-msa-sparse-attention-onto-a-3b-model-on-a-single-4090/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/i-bolted-minimax-s-msa-sparse-attention-onto-a-3b-model-on-a-single-4090/</guid><description>MSA sparse attention, retrofitted onto a 3B model on a single RTX 4090 with no training: the 28.4x compute cut holds, quality survives, plus caveats the paper skips.</description><pubDate>Thu, 25 Jun 2026 00:00:00 GMT</pubDate></item><item><title>I ran DFlash on a MacBook. The 4x is a data-center number.</title><link>https://saurabhsarkar.com/writing/i-ran-dflash-on-a-macbook-the-4x-is-a-data-center-number/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/i-ran-dflash-on-a-macbook-the-4x-is-a-data-center-number/</guid><description>I tested DFlash speculative decoding on an M5 Max and RTX 4090. The 4x headline is real, but workload and concurrency decide whether you get a speedup or a slowdown.</description><pubDate>Tue, 23 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Why your adaptive controller might be causing the defects it’s trying to prevent</title><link>https://saurabhsarkar.com/writing/why-your-adaptive-controller-might-be-causing-the-defects-it-s-trying-to-prevent/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/why-your-adaptive-controller-might-be-causing-the-defects-it-s-trying-to-prevent/</guid><description>A field note on over-cautious process control, and the margin hiding in your own sensor data.In one engagement: the force the controller watches stays low across speeds, while sidewall wear spikes at the crawl.“Slower is safer” is the most expensive assumption on your shop floor.Most adaptive controllers are built on it. When cutting force climbs, the controller slows the tool; when force eases, it speeds back up. That logic is sound for one failure mode: snapping a tool or gouging a part under</description><pubDate>Thu, 11 Jun 2026 00:00:00 GMT</pubDate></item><item><title>The Geometry of Intelligence: What Lagrange Can Teach Modern AI</title><link>https://saurabhsarkar.com/writing/the-geometry-of-intelligence-what-lagrange-can-teach-modern-ai/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/the-geometry-of-intelligence-what-lagrange-can-teach-modern-ai/</guid><description>In 1772, Joseph-Louis Lagrange solved a piece of the three-body problem that had stumped everyone before him. He didn&apos;t solve it by tracking forces. He solved it by finding the five points in the gravitational field of two large bodies where a third, smaller body can sit stably - held in place not by any single pull, but by the balance of all of them.These are the Lagrange points. They are where constraints carve stable regions out of a space that would otherwise be chaotic.That is a useful imag</description><pubDate>Fri, 01 May 2026 00:00:00 GMT</pubDate></item><item><title>The Real State of Generative AI in the Enterprise in 2025</title><link>https://saurabhsarkar.com/writing/the-real-state-of-generative-ai-in-the-enterprise/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/the-real-state-of-generative-ai-in-the-enterprise/</guid><description>A practical analysis of generative AI in the enterprise, covering pilots, production failures, governance challenges, and what actually scales.</description><pubDate>Fri, 12 Dec 2025 00:00:00 GMT</pubDate></item><item><title>AI in Finance: Transforming the Industry</title><link>https://saurabhsarkar.com/writing/how-ai-is-transforming-the-financial-industry/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/how-ai-is-transforming-the-financial-industry/</guid><description>Discover how AI transforms finance with automation, risk analysis, and customer insights. Stay ahead in the evolving financial landscape.</description><pubDate>Tue, 11 Nov 2025 00:00:00 GMT</pubDate></item><item><title>When Computation Surprises Itself: How Emergent Uncomputability Might Explain AI, Consciousness, and the Unive</title><link>https://saurabhsarkar.com/writing/when-computation-surprises-itself/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/when-computation-surprises-itself/</guid><description>Emergent unpredictability may link physics, AI, and consciousness—suggesting the universe is a self-computing system discovering itself.</description><pubDate>Wed, 05 Nov 2025 00:00:00 GMT</pubDate></item><item><title>Custom AI Solutions for Business Growth</title><link>https://saurabhsarkar.com/writing/revolutionize-operations-with-custom-ai-solutions/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/revolutionize-operations-with-custom-ai-solutions/</guid><description>Discover the benefits of custom AI solutions tailored to your business. Achieve efficiency and innovation with personalized AI strategies.</description><pubDate>Tue, 04 Nov 2025 00:00:00 GMT</pubDate></item><item><title>Fraud Detection Solutions to Protect Assets</title><link>https://saurabhsarkar.com/writing/advanced-fraud-detection-solutions-for-your-business/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/advanced-fraud-detection-solutions-for-your-business/</guid><description>Learn about cutting-edge fraud detection solutions to safeguard your business. Protect assets and build trust with innovative technology.</description><pubDate>Mon, 27 Oct 2025 00:00:00 GMT</pubDate></item><item><title>When a Few Pixels Redrawn Change Everything: The Battle Over Human Authorship in the Age of AI</title><link>https://saurabhsarkar.com/writing/intellectual-property-and-ai-generated-content-a-first-principles-view/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/intellectual-property-and-ai-generated-content-a-first-principles-view/</guid><description>Understanding Intellectual Property in the Modern AgeAt its core, intellectual property (IP) is not merely about ownership of ideas. It revolves around incentive design. Societies discovered long ago that creativity and invention flourish when individuals can reap benefits from their mental labor. However, pure ideas are non-rivalrous: once shared, they can be copied infinitely. Therefore, IP law carves out temporary monopolies on specific expressions or applications of ideas. The logic is utili</description><pubDate>Sat, 18 Oct 2025 00:00:00 GMT</pubDate></item><item><title>Animals vs Ghosts by Karpathy: A Realistic Path for Enterprise AI</title><link>https://saurabhsarkar.com/writing/animals-vs-ghosts-a-realistic-path-for-enterprise-ai/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/animals-vs-ghosts-a-realistic-path-for-enterprise-ai/</guid><description>Most AI systems today are ghosts: trained on human text, fine-tuned for alignment, and limited to observing and predicting.The future may lie in animals: AI systems that act, learn, and adapt from experience. But the path from ghost to animal is farmore complex than most realize. This article explains why enterprises should focus on perfecting &quot;ghosts with tools&quot; whilecarefully experimenting with &quot;animal instincts&quot; in controlled environments—and why this transition will take years, not months.1</description><pubDate>Tue, 07 Oct 2025 00:00:00 GMT</pubDate></item><item><title>Finance x AI: Key Trends in AI in Financial Services 2025 - Phenx Tech</title><link>https://saurabhsarkar.com/writing/finance-x-ai-key-trends-and-breakthroughs-in-2025/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/finance-x-ai-key-trends-and-breakthroughs-in-2025/</guid><description>Discover the future of AI in financial services and its impact on finance workflows. Explore key trends and breakthroughs in AI in financial services today.</description><pubDate>Tue, 30 Sep 2025 00:00:00 GMT</pubDate></item><item><title>Custom AI Solutions for Your Business</title><link>https://saurabhsarkar.com/writing/tailoring-ai-to-meet-business-needs/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/tailoring-ai-to-meet-business-needs/</guid><description>Discover the benefits of custom AI. Learn how custom AI solutions can address unique challenges in your business.</description><pubDate>Mon, 22 Sep 2025 00:00:00 GMT</pubDate></item><item><title>Build vs Buy: Navigating Software Decisions in the Age of AI</title><link>https://saurabhsarkar.com/writing/build-vs-buy-in-the-age-of-ai-the-new-executive-playbook/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/build-vs-buy-in-the-age-of-ai-the-new-executive-playbook/</guid><description>Executive SummaryFor decades, executives have debated whether to build or buy software. AI has changed the dynamics, but the fundamental principle has become clearer than ever: • Buy more often than you build. Mature, complex, regulated systems are safer to source from vendors. • Build only when it increases your competitive gap or creates a unique value proposition. That is where differentiation, margins, and long-term survival live. • Augment with AI agents to bridge gaps, integrate platforms,</description><pubDate>Thu, 28 Aug 2025 00:00:00 GMT</pubDate></item><item><title>I Read the Chat GPT-5 Prompting Playbook. Here’s My Hot Take.</title><link>https://saurabhsarkar.com/writing/i-read-the-gpt-5-prompting-playbook-here-s-my-hot-take/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/i-read-the-gpt-5-prompting-playbook-here-s-my-hot-take/</guid><description>CTO-friendly hot take on GPT-5’s Prompting Playbook: five levers that move metrics—eagerness, Responses API, verbosity, instruction hygiene, repo rules.</description><pubDate>Thu, 14 Aug 2025 00:00:00 GMT</pubDate></item><item><title>AI Trends 2025: What Karpathy’s Talk Didn’t Tell You (But You Need to Know) - Software 3.0</title><link>https://saurabhsarkar.com/writing/ai-trends-2025-software-30/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/ai-trends-2025-software-30/</guid><description>Andrej Karpathy says prompts are the new code—but that’s just the beginning. This deep dive reveals what his Software 3.0 talk missed: real-world developer challenges, second-degree insights, and the shift from coding to product judgment.</description><pubDate>Fri, 20 Jun 2025 00:00:00 GMT</pubDate></item><item><title>Apple&apos;s Illusion of Thinking paper</title><link>https://saurabhsarkar.com/writing/illusionofthinkingpaper/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/illusionofthinkingpaper/</guid><description>Apple&apos;s “Illusion of Thinking” paper reveals why reasoning models collapse—and why verification, not model size, is the next AI battleground</description><pubDate>Thu, 12 Jun 2025 00:00:00 GMT</pubDate></item><item><title>AI Trends 2025: I Read Mary Meeker’s 300-Slide Deck So You Don’t Have To</title><link>https://saurabhsarkar.com/writing/ai-trends-2025-i-read-mary-meeker-s-300-slide-deck-so-you-don-t-have-to/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/ai-trends-2025-i-read-mary-meeker-s-300-slide-deck-so-you-don-t-have-to/</guid><description>A strategic breakdown of Mary Meeker’s 2025 AI Trends report for CTOs, CIOs, and CEOs—key signals, second-order insights, and what to act on now.</description><pubDate>Tue, 03 Jun 2025 00:00:00 GMT</pubDate></item><item><title>2025’s Ultimate Guide to AI-Driven Credit Risk Models for Climate-Impacted Regions</title><link>https://saurabhsarkar.com/writing/2025-s-ultimate-guide-to-ai-driven-credit-risk-models-for-climate-impacted-regions/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/2025-s-ultimate-guide-to-ai-driven-credit-risk-models-for-climate-impacted-regions/</guid><description>Explore how AI-driven credit risk models reshape lending in climate-impacted regions. Discover 2025 trends, expert insights and tools to future-proof your portfolio.</description><pubDate>Wed, 28 May 2025 00:00:00 GMT</pubDate></item><item><title>Beyond the Buzz: Deploying Generative AI in Credit Risk decision with Precision and Accountability</title><link>https://saurabhsarkar.com/writing/generativeaiincreditrisk/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/generativeaiincreditrisk/</guid><description>Discover how top lenders are using Generative AI to speed up credit risk, reduce bias, and scale underwriting with trust and transparency.</description><pubDate>Wed, 21 May 2025 00:00:00 GMT</pubDate></item><item><title>Why AI Still Struggles with Negation and How We Can Fix It</title><link>https://saurabhsarkar.com/writing/why-ai-still-struggles-with-negation-and-how-we-can-fix-it/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/why-ai-still-struggles-with-negation-and-how-we-can-fix-it/</guid><description>AI still stumbles on negation—“not,” “no,” and absence. Discover why LLMs fail and how researchers are solving it with symbolic logic, counterfactuals &amp; more.</description><pubDate>Wed, 21 May 2025 00:00:00 GMT</pubDate></item><item><title>AI TCO Framework: Frequently Asked Questions About the True Cost of Enterprise AI</title><link>https://saurabhsarkar.com/writing/ai-tco-framework-frequently-asked-questions-about-the-true-cost-of-enterprise-ai/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/ai-tco-framework-frequently-asked-questions-about-the-true-cost-of-enterprise-ai/</guid><description>Discover the true cost of enterprise AI with our AI TCO framework. Learn what drives hidden costs, how to reduce spend, and download a free AI cost calculator. Ideal for CIOs, CTOs, and finance leaders planning sustainable AI adoption.</description><pubDate>Thu, 08 May 2025 00:00:00 GMT</pubDate></item><item><title>Will AI Agents Replace SaaS? Or Make It Invisible?</title><link>https://saurabhsarkar.com/writing/will-ai-agents-replace-saas-or-make-it-invisible/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/will-ai-agents-replace-saas-or-make-it-invisible/</guid><description>Discover how AI agents are transforming business software by automating logic, reducing backlogs, and making SaaS systems feel invisible. Learn what leading platforms like Salesforce and Microsoft are doing — and how to get started.</description><pubDate>Sun, 04 May 2025 00:00:00 GMT</pubDate></item><item><title>secure AI development, enterprise AI solutions, explainable AI, AI uncertainty</title><link>https://saurabhsarkar.com/writing/secure-enterprise-ai-expert-engineering/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/secure-enterprise-ai-expert-engineering/</guid><description>AI-generated code introduces hidden risks for enterprises. Learn how Phenx transforms AI uncertainty into secure, explainable, enterprise-ready AI systems through expert engineering and client collaboration.</description><pubDate>Tue, 29 Apr 2025 00:00:00 GMT</pubDate></item><item><title>From Tool Soup to Protocol Power: Why MCP Server Might Just Be the API Revolution AI&apos;s Been Waiting For</title><link>https://saurabhsarkar.com/writing/from-tool-soup-to-protocol-power-mcp-server-might-just-be-the-api-revolution-ai-s-been-waiting-for/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/from-tool-soup-to-protocol-power-mcp-server-might-just-be-the-api-revolution-ai-s-been-waiting-for/</guid><description>Discover how the Model Context Protocol (MCP) server revolutionizes AI tool integration. Learn how businesses and developers can use MCP to build modular, scalable, and intelligent AI systems—without prompt chaos.</description><pubDate>Wed, 09 Apr 2025 00:00:00 GMT</pubDate></item><item><title>The Pizza Paradigm: A Manifesto on the Art and Philosophy of AI Prompting</title><link>https://saurabhsarkar.com/writing/the-pizza-paradigm-a-manifesto-on-the-art-and-philosophy-of-ai-prompting/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/the-pizza-paradigm-a-manifesto-on-the-art-and-philosophy-of-ai-prompting/</guid><description>Four ways to structure AI prompts, explained through pizza, music, visual art, and the creative process.</description><pubDate>Sat, 08 Mar 2025 00:00:00 GMT</pubDate></item><item><title>Why LLMs Are the Future of IT Automation: Beyond Simple Rule-Based Systems</title><link>https://saurabhsarkar.com/writing/why-llms-are-the-future-of-it-automation-beyond-simple-rule-based-systems/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/why-llms-are-the-future-of-it-automation-beyond-simple-rule-based-systems/</guid><description>Discover why LLMs and Future of IT Automation are game-changers for IT systems. Learn how LLMs enhance operations beyond rule-based limits.</description><pubDate>Sat, 22 Feb 2025 00:00:00 GMT</pubDate></item><item><title>The Future of AI Agent Frameworks: Trends, Predictions, and Business Opportunities</title><link>https://saurabhsarkar.com/writing/the-future-of-ai-agent-frameworks-trends-predictions-and-business-opportunities/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/the-future-of-ai-agent-frameworks-trends-predictions-and-business-opportunities/</guid><description>Discover the future of AI agent frameworks with insights into hybrid systems, low-code platforms, explainable AI, and edge computing. Learn how these trends are transforming industries and creating smarter, scalable, and ethical AI solutions.</description><pubDate>Tue, 14 Jan 2025 00:00:00 GMT</pubDate></item><item><title>How AI agents will Automate Enterprise workflow?</title><link>https://saurabhsarkar.com/writing/how-ai-agents-will-automate-enterprise-workflow/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/how-ai-agents-will-automate-enterprise-workflow/</guid><description>Discover how AI agents are revolutionizing enterprise workflows by automating repetitive tasks, enhancing decision-making, and orchestrating complex operations. Learn about their impact, real-world success stories, and how they’re shaping the future of work.</description><pubDate>Tue, 31 Dec 2024 00:00:00 GMT</pubDate></item><item><title>AI vs. Agents: Understanding Their Differences and Collaborative Roles</title><link>https://saurabhsarkar.com/writing/https-www-phenx-io-post-ai-vs-agents-differences-collaboration/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/https-www-phenx-io-post-ai-vs-agents-differences-collaboration/</guid><description>Explore the distinctions and collaborative dynamics between Artificial Intelligence (AI) vs agents. Learn how thinkers and doers in technology shape modern innovations.</description><pubDate>Tue, 31 Dec 2024 00:00:00 GMT</pubDate></item><item><title>Agents Unleashed: The Top 5 AI Agent Types Powering the Future</title><link>https://saurabhsarkar.com/writing/agents-unleashed-the-top-5-ai-types-powering-the-future/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/agents-unleashed-the-top-5-ai-types-powering-the-future/</guid><description>Discover the top 5 AI agents revolutionizing technology and business. Learn about their types, real-world applications, and why executives should care. Explore how AI agents like learning and utility-based systems drive innovation in industries from robotics to finance.</description><pubDate>Thu, 26 Dec 2024 00:00:00 GMT</pubDate></item><item><title>Harnessing the Creative Power of AI Hallucinations: A New Frontier for Innovation</title><link>https://saurabhsarkar.com/writing/harnessing-the-creative-power-of-ai-hallucinations-a-new-frontier-for-innovation/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/harnessing-the-creative-power-of-ai-hallucinations-a-new-frontier-for-innovation/</guid><description>Discover how to turn AI hallucinations into a powerful tool for innovation. Learn strategies to control and harness AI&apos;s creative potential for breakthroughs in science, technology, and business.</description><pubDate>Thu, 26 Dec 2024 00:00:00 GMT</pubDate></item><item><title>The Art of AI Traffic Control: Mastering LLM Routing</title><link>https://saurabhsarkar.com/writing/the-art-of-ai-traffic-control-mastering-llm-routing/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/the-art-of-ai-traffic-control-mastering-llm-routing/</guid><description>Discover how LLM routing is transforming industries by optimizing AI-powered workflows. Learn real-world use cases, from customer support to healthcare, where task-specific models improve efficiency, reduce costs, and enhance performance.</description><pubDate>Mon, 25 Nov 2024 00:00:00 GMT</pubDate></item><item><title>Can AIs Debate Their Way to the Truth? Exploring Experimental AI Debates with a Guidance-Focused Judge</title><link>https://saurabhsarkar.com/writing/can-ais-debate-their-way-to-the-truth-exploring-experimental-ai-debates-with-a-guidance-focused-jud/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/can-ais-debate-their-way-to-the-truth-exploring-experimental-ai-debates-with-a-guidance-focused-jud/</guid><description>This blog post delves into a fascinating thought experiment where two AI models engage in a structured debate over the nature of a simulated surface—flat or round. Using a debate framework guided by a sophisticated, evidence-focused judge, we explore how AIs can be trained to prioritize truth over rhetorical skill. The post examines the importance of rigorous experimentation, the role of guidance-based judgment, and the broader implications of debate-driven truth-seeking for AI alignment.</description><pubDate>Fri, 08 Nov 2024 00:00:00 GMT</pubDate></item><item><title>How Can Fuzzy Matching and NER Solve Complex Business Data Challenges?</title><link>https://saurabhsarkar.com/writing/how-can-fuzzy-matching-and-ner-solve-complex-business-data-challenges/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/how-can-fuzzy-matching-and-ner-solve-complex-business-data-challenges/</guid><description>ransform unstructured business data with fuzzy matching and NER. This guide reveals how Python can help solve complex data inconsistencies for improved accuracy in machine learning projects.</description><pubDate>Wed, 30 Oct 2024 00:00:00 GMT</pubDate></item><item><title>Behind the Curtain: Why AI Explainability Matters Now More Than Ever</title><link>https://saurabhsarkar.com/writing/behind-the-curtain-why-ai-explainability-matters-now-more-than-ever/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/behind-the-curtain-why-ai-explainability-matters-now-more-than-ever/</guid><description>Explore the importance of explainable AI in public sector applications, with insights on balancing complexity and transparency. Discover practical templates, stakeholder-focused explanations, and ethical considerations, guiding responsible AI development for trust and accountability.</description><pubDate>Tue, 29 Oct 2024 00:00:00 GMT</pubDate></item><item><title>What Happens When AI Faces Uncertainty? Unlocking the Power of Probabilistic Reasoning in Language Models</title><link>https://saurabhsarkar.com/writing/what-happens-when-ai-faces-uncertainty-unlocking-the-power-of-probabilistic-reasoning-in-language-m/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/what-happens-when-ai-faces-uncertainty-unlocking-the-power-of-probabilistic-reasoning-in-language-m/</guid><description>Discover the power of probabilistic reasoning in AI. Learn how Language Models are evolving to make flexible, nuanced decisions in uncertain environments.</description><pubDate>Wed, 23 Oct 2024 00:00:00 GMT</pubDate></item><item><title>The Power of RPA + AI for Heavy Industries: Automating for Efficiency, Safety, and Predictive Insights</title><link>https://saurabhsarkar.com/writing/the-power-of-rpa-ai-for-heavy-industries-automating-for-efficiency-safety-and-predictive-insigh/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/the-power-of-rpa-ai-for-heavy-industries-automating-for-efficiency-safety-and-predictive-insigh/</guid><description>Discover how RPA and AI are transforming heavy industries like Oil &amp; Gas, Manufacturing, and Construction. Learn how automation boosts efficiency, enhances safety, and delivers predictive insights for smarter decision-making and optimized operations.</description><pubDate>Tue, 22 Oct 2024 00:00:00 GMT</pubDate></item><item><title>LoLCATs: Demystifying Linearized Attention in Large Language Models</title><link>https://saurabhsarkar.com/writing/lolcats-demystifying-linearized-attention-in-large-language-models/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/lolcats-demystifying-linearized-attention-in-large-language-models/</guid><description>Discover how LoLCATs (Learnable Linearized Attention Transformers) revolutionize large language models by simplifying the attention mechanism from quadratic to linear complexity. This blog explores the use of learnable linear attention, low-rank adaptation (LoRA), and layer-wise optimization to make LLMs more efficient, scalable, and accessible. Learn how LoLCATs enable models to handle larger sequences with reduced computational costs, while maintaining performance.</description><pubDate>Wed, 16 Oct 2024 00:00:00 GMT</pubDate></item><item><title>Agents on Rails: The Smartest Way to Keep AI Under Control and On Task</title><link>https://saurabhsarkar.com/writing/agents-on-rails-the-smartest-way-to-keep-ai-under-control-and-on-task/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/agents-on-rails-the-smartest-way-to-keep-ai-under-control-and-on-task/</guid><description>&quot;Agents on rails&quot; refers to a structured framework for autonomous AI agents, where their actions, decisions, and processes are guided by predefined pathways or &quot;rails.&quot; This approach limits the unpredictability of agents by constraining them to follow specific rules or workflows, ensuring that they remain aligned with business objectives while carrying out complex tasks. This architecture balances flexibility with control, helping businesses automate processes efficiently.</description><pubDate>Fri, 27 Sep 2024 00:00:00 GMT</pubDate></item><item><title>Navigating the AI Seas: A CIO’s Guide to AGI Safety and Alignment</title><link>https://saurabhsarkar.com/writing/navigating-the-ai-seas-a-cio-s-guide-to-agi-safety-and-alignment/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/navigating-the-ai-seas-a-cio-s-guide-to-agi-safety-and-alignment/</guid><description>Explore how CIOs can navigate the challenges of Artificial General Intelligence (AGI) with a focus on safety and alignment. Discover actionable strategies to harness AGI&apos;s potential while mitigating risks, set in a near-future scenario where AGI is on the horizon. AI Safety is key for enterprise AI journey.</description><pubDate>Mon, 26 Aug 2024 00:00:00 GMT</pubDate></item><item><title>Streamlining AI: The Essentials of Quantization and Model Distillation for Large Language Models</title><link>https://saurabhsarkar.com/writing/streamlining-ai-the-essentials-of-quantization-and-model-distillation-for-large-language-models/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/streamlining-ai-the-essentials-of-quantization-and-model-distillation-for-large-language-models/</guid><description>With Large Language Models (LLMs) like OpenAI&apos;s GPT becoming increasingly prevalent, optimizing them for better performance and lower resource consumption has become a top priority. Two techniques stand out in this optimization landscape: quantization and model distillation.</description><pubDate>Thu, 27 Jun 2024 00:00:00 GMT</pubDate></item><item><title>Deciphering the Titans of Thought: General AI (AGI) vs. Strong AI (ASI)</title><link>https://saurabhsarkar.com/writing/deciphering-the-titans-of-thought-artificial-general-ai-vs-artificial-strong-ai/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/deciphering-the-titans-of-thought-artificial-general-ai-vs-artificial-strong-ai/</guid><description>Unravel the enigmatic world of AI with our guide to Artificial General AI vs. Artificial Strong AI. Explore the cosmos of AGI and ASI with us!</description><pubDate>Thu, 13 Jun 2024 00:00:00 GMT</pubDate></item><item><title>AI for Business Automation: Essential Strategies for C-Suite Executives</title><link>https://saurabhsarkar.com/writing/ai-for-business-automation-essential-strategies-for-c-suite-executives/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/ai-for-business-automation-essential-strategies-for-c-suite-executives/</guid><description>By harnessing the capabilities of AI, Business can automate mundane tasks, freeing up valuable resources to focus on strategic initiatives.</description><pubDate>Fri, 07 Jun 2024 00:00:00 GMT</pubDate></item><item><title>Behind the Curtain: Quirks and Perks of Enterprise AI in Finance - Prompt Engineering, RAG, and More</title><link>https://saurabhsarkar.com/writing/behind-the-curtain-quirks-and-perks-of-enterprise-ai-in-finance-prompt-engineering-rag-and-more/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/behind-the-curtain-quirks-and-perks-of-enterprise-ai-in-finance-prompt-engineering-rag-and-more/</guid><description>Unleash the full potential of Enterprise AI in finance through the art of prompt engineering and the transformative power of Retrieval-Augmented Generation (RAG).</description><pubDate>Thu, 06 Jun 2024 00:00:00 GMT</pubDate></item><item><title>Generative AI vs. Applied AI in Business</title><link>https://saurabhsarkar.com/writing/generative-ai-vs-applied-ai-in-business/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/generative-ai-vs-applied-ai-in-business/</guid><description>Both Generative AI and Applied AI are integral parts of the broader Business AI landscape but serve different purposes. Generative AI focuses on creating new content and ideas, while Applied AI is about applying AI technologies to improve processes, enhance decision-making, or automate tasks across various industries.</description><pubDate>Thu, 06 Jun 2024 00:00:00 GMT</pubDate></item><item><title>Adventures in AI: Duel of the Digital Minds - Exploring the RPG World with Two Unique LLM Architectures, Act 2</title><link>https://saurabhsarkar.com/writing/adventures-in-ai-duel-of-the-digital-minds-exploring-the-rpg-world-with-two-unique-llm-architectu/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/adventures-in-ai-duel-of-the-digital-minds-exploring-the-rpg-world-with-two-unique-llm-architectu/</guid><description>Building on the foundation in &quot;Navigating New Realms: Open-Source LLMs in RPG Environments,&quot; our follow-up exploration in LLM innovative application series, pushes the envelope further by comparing the outcomes of two different models in controlling a player within a virtual RPG world. The first model, a 7B Dolphin2.2-Mistral, demonstrated impressive capabilities in navigating and strategizing within a complex, dynamically rendered RPG landscape. This model, built on interpreting textual descrip</description><pubDate>Sat, 02 Mar 2024 00:00:00 GMT</pubDate></item><item><title>Navigating New Realms: Open-Source LLMs in RPG Environments</title><link>https://saurabhsarkar.com/writing/navigating-new-realms-open-source-llms-in-rpg-environments/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/navigating-new-realms-open-source-llms-in-rpg-environments/</guid><description>The crossroads of AI and gaming are more than just a breeding ground for innovation; they are testing grounds for the future. Our latest venture pushes this boundary by integrating Large Language Models (LLMs) into the heart of role-playing games (RPGs). This project, rooted in open-source and community collaboration, puts LLMs to the test in a dynamic RPG universe. Here, LLMs face a world that changes with every decision, challenging them to devise strategies in real-time. But this is about mor</description><pubDate>Wed, 21 Feb 2024 00:00:00 GMT</pubDate></item><item><title>Security of AI Models: Navigating Emerging Threats and Solutions</title><link>https://saurabhsarkar.com/writing/security-of-ai-models-navigating-emerging-threats-and-solutions/</link><guid isPermaLink="true">https://saurabhsarkar.com/writing/security-of-ai-models-navigating-emerging-threats-and-solutions/</guid><description>Executive SummaryThis blog post provides a comprehensive exploration of the current security challenges faced in the field of Artificial Intelligence (AI) . This document serves as a critical guide for understanding and addressing the various types of attacks that AI models are susceptible to, including data poisoning, membership inference attacks, model extraction attacks, and the practice of fairwashing. The white paper aims to educate and inform a wide range of audiences, from AI professional</description><pubDate>Mon, 05 Feb 2024 00:00:00 GMT</pubDate></item></channel></rss>