LLM

  • 19 July 2026
  • img/cover.webp

Adding Langfuse Observability to Hermes Agent (Self-Hosted, the Hard Parts Included)

How do you add observability to Hermes Agent with self-hosted Langfuse? A walkthrough of enabling the bundled plugin, the SDK gotcha that silently kills tracing, per-model cost tracking on Coolify, and what the traces actually look like.

Read more 
  • 20 June 2026
  • cover.webp

The Trends of Engineering in AI: Prompt, Context, Harness, Loop, and Evaluation

What’s the difference between prompt engineering, context engineering, harness engineering, loop engineering, and evaluation engineering? A side-by-side comparison with diagrams for AI engineers building production systems.

Read more 
  • 29 May 2026
  • cover.webp

My Subagents Lied to Me: What Happened When I Let AI Research Autonomously

How do you know if an AI agent is telling the truth? When my research agents fabricated 11 model names and 4 fake paper titles, I learned the hard way that autonomous AI research requires independent verification.

Read more 
  • 19 February 2026
  • cover-image.webp

Miss One Weekend, Fall Behind One Month

I went on a weekend trip. I came back to a new Claude, a new GPT and an existential crisis. The pace of AI is no longer monthly. It’s weekly.

Read more 
  • 25 January 2026
  • ai_in_ai-cover.webp

AI in AI: What I Learned Analysing AI in the Adult Industry

I went down the rabbit hole of AI in the adult industry. I expected simple chatbots; I found a sophisticated engineering stack pushing the boundaries of edge computing, privacy and opensource AI.

Read more 
  • 27 June 2025
  • observability-3_achieving_ai_observability_light.webp

Advancing AI Observability: From Metrics to Meaningful Insights

How do you monitor AI systems in production? Practical strategies for instrumenting LLM applications with OpenTelemetry, tracking token costs, latency and quality metrics.

Read more 
  • 11 May 2025
  • llm-evals-light-1.webp

The Hidden Cost of LLM-as-a-Judge: When More Evaluation Means Less Value

What are the hidden costs of using LLM-as-a-Judge? Learn about common biases, failure patterns and smarter evaluation strategies for LLM applications.

Read more 
  • 12 January 2024
  • llmops-introduction.webp

LLMOps: Introduction

What is LLMOps? An introduction to Large Language Model Operations covering the lifecycle from DevOps to MLOps to LLMOps, with key differences and tooling.

Read more 
  • 13 September 2023
  • llm-history.webp

Large Language Models History

What is the history of Large Language Models? From rule-based systems to transformers to GPT-4: understand the evolution of LLMs and their impact on NLP.

Read more 
  • 17 March 2023
  • prompt-engineering.webp

Prompt Engineering

What is prompt engineering and how does it work? Learn zero-shot, few-shot, chain-of-thought, generated knowledge and ReAct prompting techniques for LLMs.

Read more 
  • 14 March 2023
  • access-gpt4.webp

How to access GPT-4

How to access GPT-4 from OpenAI.

Read more 
  • 08 February 2023
  • chatgpt-alternative.webp

ChatGPT Alternatives

ChatGPT Alternatives: ChatGPT, Bard, and other alternatives.

Read more