<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0">
  <channel>
    <title>OpenTeams Engineering Blog</title>
    <link>https://openteams.com/engineering-blog/</link>
    <description>Benchmarks, systems deep dives, and hard-won lessons from OpenTeams engineers.</description>
    <item>
      <title>Reliable Visual Regression Testing for Humans and Coding Agents</title>
      <link>https://openteams.com/engineering-blog/visual-regression-testing-jupyterlab/</link>
      <guid isPermaLink="true">https://openteams.com/engineering-blog/visual-regression-testing-jupyterlab/</guid>
      <pubDate>Tue, 06 Oct 2026 14:00:12 GMT</pubDate>
      <description>See how JupyterLab made visual regression testing reproducible on Linux machines, cut CI from 55 to 15 minutes and let contributors update snapshots.</description>
    </item>
    <item>
      <title>Your AI Agent Ignored AGENTS.md. Your Linter Won't Let It.</title>
      <link>https://openteams.com/engineering-blog/lint-rules-for-ai-agents/</link>
      <guid isPermaLink="true">https://openteams.com/engineering-blog/lint-rules-for-ai-agents/</guid>
      <pubDate>Thu, 01 Oct 2026 07:26:51 GMT</pubDate>
      <description>Use lint rules for AI agents to enforce project conventions. Custom ESLint rules like @jupyter/eslint-plugin give agents clear errors they can fix on their own.</description>
    </item>
    <item>
      <title>Where Does Jev Fit in a Software Supply Chain? We Benchmarked It Against the Open Alternatives</title>
      <link>https://openteams.com/engineering-blog/jev-bench-package-curation/</link>
      <guid isPermaLink="true">https://openteams.com/engineering-blog/jev-bench-package-curation/</guid>
      <pubDate>Tue, 29 Sep 2026 08:59:43 GMT</pubDate>
      <description>We benchmarked TypeSafe's Jev, Laya, CLM-8B and Claude Haiku on five package-curation triage tasks. A fine-tuned 421M model matched Jev at a tenth of the latency.</description>
    </item>
    <item>
      <title>Awesome Jupyter AI: A Map of 100+ Jupyter Extensions for AI</title>
      <link>https://openteams.com/engineering-blog/awesome-jupyter-ai-extensions/</link>
      <guid isPermaLink="true">https://openteams.com/engineering-blog/awesome-jupyter-ai-extensions/</guid>
      <pubDate>Fri, 25 Sep 2026 12:52:56 GMT</pubDate>
      <description>Find the right Jupyter AI extension: a curated map of 100+ chat panels, inline completers, agent bridges and MCP servers for JupyterLab and Notebook 7.</description>
    </item>
    <item>
      <title>MiniCPM5-2B quantization report</title>
      <link>https://openteams.com/engineering-blog/minicpm5-2b-quantization-report/</link>
      <guid isPermaLink="true">https://openteams.com/engineering-blog/minicpm5-2b-quantization-report/</guid>
      <pubDate>Wed, 23 Sep 2026 01:18:00 GMT</pubDate>
      <description>MiniCPM5-2B quantization report: the best GGUF weights and K/V cache quants on llama.cpp and BeeLlama.cpp, squeezing a SOTA model into 3 GiB RAM.</description>
    </item>
    <item>
      <title>I Asked LLMs to Review Another LLM. They Still Got It Wrong</title>
      <link>https://openteams.com/engineering-blog/llm-review-reliability/</link>
      <guid isPermaLink="true">https://openteams.com/engineering-blog/llm-review-reliability/</guid>
      <pubDate>Fri, 18 Sep 2026 02:24:06 GMT</pubDate>
      <description>I tested LLM reviewer models on meeting summaries to see when they catch hallucinations, when they delete true claims, and how to choose one.</description>
    </item>
    <item>
      <title>LLMs: Intelligence vs. cost</title>
      <link>https://openteams.com/engineering-blog/intelligence-vs-cost/</link>
      <guid isPermaLink="true">https://openteams.com/engineering-blog/intelligence-vs-cost/</guid>
      <pubDate>Wed, 02 Sep 2026 01:29:09 GMT</pubDate>
      <description>LLM intelligence vs cost: logarithmic scales in plots hide the cost chasm between frontier and cheap models; actual OpenRouter prices alter the Pareto frontier.</description>
    </item>
    <item>
      <title>pixi on UBI-micro: A Safer, Smaller Multi-Stage Container Build</title>
      <link>https://openteams.com/engineering-blog/pixi-ubi-micro-containers/</link>
      <guid isPermaLink="true">https://openteams.com/engineering-blog/pixi-ubi-micro-containers/</guid>
      <pubDate>Fri, 24 Jul 2026 12:21:21 GMT</pubDate>
      <description>How a UBI-micro multi-stage build shrinks a pixi container's attack surface and CVE count, with side-by-side Trivy and Grype scans on UBI full, minimal, and micro. </description>
    </item>
    <item>
      <title>The Best Code Review Says Less</title>
      <link>https://openteams.com/engineering-blog/best-code-reviews-say-less/</link>
      <guid isPermaLink="true">https://openteams.com/engineering-blog/best-code-reviews-say-less/</guid>
      <pubDate>Fri, 24 Jul 2026 12:06:01 GMT</pubDate>
      <description>Far too many AI code review comments are noise. Pair a human editor with AI, surface only the findings that matter, and encode that discipline in a reusable skill.</description>
    </item>
    <item>
      <title>You Don't Drive AI. You Ride It.</title>
      <link>https://openteams.com/engineering-blog/you-dont-drive-ai-you-ride-it/</link>
      <guid isPermaLink="true">https://openteams.com/engineering-blog/you-dont-drive-ai-you-ride-it/</guid>
      <pubDate>Thu, 09 Jul 2026 21:37:12 GMT</pubDate>
      <description>AI is stochastic, not deterministic. You don't drive it, you ride it: make documents the memory, use the variation to prototype, and verify every public claim.</description>
    </item>
    <item>
      <title>Who Is Your Code's Audience? An Engineering Team Talks AI-Written Code</title>
      <link>https://openteams.com/engineering-blog/who-is-your-codes-audience/</link>
      <guid isPermaLink="true">https://openteams.com/engineering-blog/who-is-your-codes-audience/</guid>
      <pubDate>Thu, 09 Jul 2026 21:36:59 GMT</pubDate>
      <description>Learn why your code's audience decides how much readability matters, as an engineering team debates AI-written code, review, and accountability.</description>
    </item>
    <item>
      <title>The tp_as_number Slot and Binary Operation Dispatch in CPython</title>
      <link>https://openteams.com/engineering-blog/tp-as-number-slot-and-binary-dispatch/</link>
      <guid isPermaLink="true">https://openteams.com/engineering-blog/tp-as-number-slot-and-binary-dispatch/</guid>
      <pubDate>Mon, 22 Jun 2026 14:01:40 GMT</pubDate>
      <description>How CPython dispatches binary operations via tp_as_number slots, and how TorchDynamo mirrors that model to improve correctness and reduce ad-hoc special cases.</description>
    </item>
    <item>
      <title>PDF Table Extraction: Docling vs Marker vs LlamaParse Compared</title>
      <link>https://openteams.com/engineering-blog/docling-vs-marker-vs-llamaparse/</link>
      <guid isPermaLink="true">https://openteams.com/engineering-blog/docling-vs-marker-vs-llamaparse/</guid>
      <pubDate>Wed, 20 May 2026 03:52:48 GMT</pubDate>
      <description>Compare three Python tools for PDF table extraction: Docling, Marker, and LlamaParse. Learn which handles merged cells and multi-level headers best.</description>
    </item>
    <item>
      <title>Sandboxing Code Mode for Local LLM Agents</title>
      <link>https://openteams.com/engineering-blog/code-mode-sandboxing-local-llms/</link>
      <guid isPermaLink="true">https://openteams.com/engineering-blog/code-mode-sandboxing-local-llms/</guid>
      <pubDate>Tue, 05 May 2026 03:01:31 GMT</pubDate>
      <description>Code mode can make local LLM agents more practical, but executing model-written code brings sandboxing back into the architecture.</description>
    </item>
    <item>
      <title>Plugin Playground AI Integration for Faster Plugin Prototyping</title>
      <link>https://openteams.com/engineering-blog/plugin-playground-ai-integration/</link>
      <guid isPermaLink="true">https://openteams.com/engineering-blog/plugin-playground-ai-integration/</guid>
      <pubDate>Sat, 25 Apr 2026 02:23:05 GMT</pubDate>
      <description>Learn how we integrated AI into Plugin Playground to help you create, edit, test, package, and share JupyterLab plugins faster in JupyterLite and Binder.</description>
    </item>
    <item>
      <title>Slow Down — Simple Lessons for Guiding AI and Shipping Better Code</title>
      <link>https://openteams.com/engineering-blog/slow-down-ship-better-code/</link>
      <guid isPermaLink="true">https://openteams.com/engineering-blog/slow-down-ship-better-code/</guid>
      <pubDate>Thu, 23 Apr 2026 08:32:20 GMT</pubDate>
      <description>Practical lessons for shipping better code, staying in control, keeping your skills sharp, and getting real value from AI coding tools without losing yourself in the hype.</description>
    </item>
    <item>
      <title>We Benchmarked 6 Python Package Managers on a Real ML Project. Here's What We Found.</title>
      <link>https://openteams.com/engineering-blog/benchmark-python-package-managers/</link>
      <guid isPermaLink="true">https://openteams.com/engineering-blog/benchmark-python-package-managers/</guid>
      <pubDate>Wed, 22 Apr 2026 13:07:27 GMT</pubDate>
      <description>Head-to-head benchmark of pixi, uv, conda, mamba, pip, and poetry on a real ML/AI project with 25+ mixed conda-forge and PyPI dependencies.</description>
    </item>
    <item>
      <title>What I Learned Making a Local LLM Do Real Work</title>
      <link>https://openteams.com/engineering-blog/what-i-learned-making-a-local-llm-do-real-work/</link>
      <guid isPermaLink="true">https://openteams.com/engineering-blog/what-i-learned-making-a-local-llm-do-real-work/</guid>
      <pubDate>Wed, 22 Apr 2026 04:05:08 GMT</pubDate>
      <description>Learn what makes a local LLM agent reliable enough for real work: evals, deterministic Python logic, and why a better model beats engineering.</description>
    </item>
    <item>
      <title>From Skill to Agent: When a Text File Isn't Enough</title>
      <link>https://openteams.com/engineering-blog/from-skill-to-agent/</link>
      <guid isPermaLink="true">https://openteams.com/engineering-blog/from-skill-to-agent/</guid>
      <pubDate>Wed, 22 Apr 2026 04:05:06 GMT</pubDate>
      <description>When does a Claude Code skill stop being enough? See why credential security pushes real workflows toward proper agent architectures.</description>
    </item>
  </channel>
</rss>