Paul SerbanSoftware Engineer
  • PORTFOLIO
  • BLOG

#Retrieval Augmented Generation

Posts

  • Architecting Production LLM Systems: How AI Gateways, Orchestration, RAG, and Observability Fit TogetherA practical reference architecture connecting OpenRouter, LangChain/LlamaIndex, vector search, structured outputs, and Langfuse into one coherent AI engineering stack

    How AI gateways, LangChain/LlamaIndex orchestration, vector search, structured outputs, and Langfuse observability combine into a production LLM architecture.

    • #AI Engineering
    • #LLM Architecture
    • #Retrieval Augmented Generation
    • #AI Gateway
    • #LangChain
  • AI Embeddings Explained: Foundations, Fundamentals, and When to Use ThemWhat embeddings actually are, how they're trained, and how to use them correctly in semantic search, RAG, and clustering systems

    A practical guide to AI embeddings: what they are, how similarity search works, when to use them, and how to avoid the pitfalls that quietly break RAG systems

    • #AI And Machine Learning
    • #Vector Embeddings
    • #Semantic Search
    • #Retrieval Augmented Generation
    • #Vector Databases
View all posts
  • LinkedInLinkedIn
  • GitHubGitHub
  • HackerRankHackerRank
  • LeetCodeLeetCode
  • EmailEmail
  • Portfolio
  • My Projects
  • Coursework
  • Blog
  • Posts
  • Snippets
  • Book Notes
  • Cookie Settings
  • Cookie Policy
2026 © Paul Serban. All rights reserved.www.paulserban.eu | Sitemap