Skip to content
KG.
  • Hire me, or see what I ship.

    • ProjectsCase studies & shipped work
    • ServicesWork with me
    • SponsorSponsor the blog
  • Essays and references on AI engineering.

    • BlogEssays on AI engineering
    • Learning PathsGuided curricula
    • Topic PillarsDeep-dive hubs
    • GlossaryAI & dev terms, defined
    • CheatsheetsQuick references
    • ComparisonsX vs Y, decided
  • Things to click, play, and take apart.

    • Tools28 dev & AI utilities
    • Games16 browser games
    • DemosInteractive explainers
    • ChallengesDaily coding puzzles
  • Me

    • AboutWho I am
    • UsesMy gear & setup
    • ResumeCV (PDF)

    Shelf

    • BookshelfBooks I recommend
    • Reading ListWhat I'm reading
  • Let's talk
  1. Home
  2. ›
  3. Blog
  4. ›
  5. #llm-infrastructure

#llm-infrastructure

3 posts tagged with #llm-infrastructure

Every article below is hand-written, technically reviewed, and focused on llm-infrastructure. Posts cover real-world architecture decisions, code-level implementation patterns, and trade-offs you'll only discover after shipping production systems.

Pinecone vs Weaviate 2026: Which Vector DB Actually Wins? AI and Machine Learning

Pinecone vs Weaviate 2026: Which Vector DB Actually Wins?

Pinecone wins for teams that need zero-ops managed infrastructure and fast time-to-production. Weaviate wins for teams that want open-source flexibility, hybrid search, and full data sovereignty.

May 10, 2026 11 min read
Read more
Qdrant vs Chroma 2026: Which Open-Source Vector DB Wins for RAG? AI and Machine Learning

Qdrant vs Chroma 2026: Which Open-Source Vector DB Wins for RAG?

Qdrant wins for production RAG at scale; Chroma wins for local prototyping and developer speed. Here's the full breakdown to help you choose the right vector database before you're locked in.

May 10, 2026 12 min read
Read more
Claude Haiku 4.5 vs Llama 3 70B Local: Cost & Quality in 2026 AI and Machine Learning

Claude Haiku 4.5 vs Llama 3 70B Local: Cost & Quality in 2026

Claude Haiku 4.5 wins for zero-ops, high-volume API workloads; Llama 3 70B wins for privacy-first, cost-at-scale self-hosted deployments. Here's the full breakdown.

May 10, 2026 9 min read
Read more

Content

  • Blog
  • Topic Pillars
  • Learning Paths
  • Games
  • Demos

Resources

  • Tools
  • Glossary
  • Cheatsheets
  • Comparisons

About Me

  • About
  • Uses
  • Bookshelf
  • Reading List

Meta

  • Subscribe
  • Changelog
  • Sitemap
  • Privacy
  • Terms
  • RSS
KG

Building intelligent systems. Still chasing those sour icecreams.

Made with coffee and curiosity in Toronto. 2026.