RSSAmplifier

Blog

Hugo Nogueira

AI-forward CPTO building the future of enterprise software. 20+ years leading product x engineering x design teams. Currently shipping the most modern GRC platform at Complyance.

hugo.imRSS feed ↗49 posts

Latest posts

What 639,000 Execution Steps Taught Me About How AI Agents Really Fail

I applied the MAST failure taxonomy to 639,000 execution steps from AI agents running in production for five months. My first headline finding turned out to be an infrastructure bug masquerading as agent behavior. This post is about what agents actually fail at in production, and the discipline it takes to not fool yourself with production data.

O Que 639.000 Passos de Execução Me Ensinaram Sobre Como Agentes de IA Realmente Falham

Apliquei a taxonomia de falhas MAST a 639.000 passos de execução de agentes de IA rodando em produção por cinco meses. Meu primeiro headline acabou sendo um bug de infraestrutura disfarçado de comportamento do agente. Este post é sobre o que os agentes de fato falham em produção, e sobre a disciplina necessária para não enganar a si mesmo com dados de produção.

LHC v0.2: A Benchmark for Long-Horizon Agent Coherence (and the Methodology That Got It Honest)

I just published LHC v0.2, an open benchmark for long-horizon coherence in 8B-class agent models, plus a deterministic parser baseline that puts a useful floor on what fine-tuning is worth for structured-state tasks. This post explains what they're for, how to use them, and the methodology arc that produced them across five rounds of external review.

LHC v0.2: Um Benchmark para Coerência de Longo Horizonte em Agentes (e a Metodologia que Tornou os Resultados Honestos)

Acabei de publicar o LHC v0.2, um benchmark aberto para coerência de longo horizonte em modelos de agentes da classe 8B, mais um baseline de parser determinístico que coloca um piso útil sobre o que fine-tuning vale para tarefas de estado estruturado. Este post explica para que servem, como usá-los, e o arco metodológico que os produziu ao longo de cinco rodadas de revisão externa.

We're Mistaking the Bootstrap Phase for the Future of AI Agents

The self-hosted AI agent movement is real and important. But we are confusing a bootstrap phase with a destination architecture. The long-term future of agents will be defined by platforms that make them reliable, governable, and operationally boring.

Estamos Confundindo a Fase de Bootstrap com o Futuro dos Agentes de IA

O movimento de agentes de IA self-hosted e real e importante. Mas estamos confundindo uma fase de bootstrap com uma arquitetura de destino. O futuro de longo prazo dos agentes sera definido por plataformas que os tornem confiaveis, governaveis e operacionalmente entediantes.

Execution Is Cheap. Judgment Isn't: AI Agents and the Collapse of the CTO/CPO Divide

When execution becomes abundant through AI agents, judgment becomes the bottleneck. The traditional separation between technical and product leadership breaks down, creating space for the CPTO role.

Execução Ficou Barata. Julgamento, Não: Agentes de IA e o Fim da Divisão CTO/CPO

Quando a execução se torna abundante graças aos agentes de IA, o julgamento vira o gargalo. A separação tradicional entre liderança técnica e de produto começa a não fazer mais sentido.

The OWASP Top 10 for AI Agents - Security in the Age of Autonomy

OWASP just published their first Top 10 for Agentic Applications. Here's what every agent builder needs to know about the new attack surfaces, why traditional security fails, and how the orchestration layer becomes the new security boundary.

OWASP Top 10 para Agentes de IA - Segurança na Era da Autonomia

O OWASP acaba de publicar seu primeiro Top 10 para Aplicações Agênticas. Aqui está o que todo construtor de agentes precisa saber sobre as novas superfícies de ataque, por que a segurança tradicional falha e como a camada de orquestração se torna o novo limite de segurança.

From OpenClaw's Chaos to OpenAI's Frontier: The Agent Infrastructure Reckoning

OpenClaw went from viral sensation to security nightmare. Today, OpenAI launched Frontier. Two sides of the same story: autonomous agents are going mainstream, and the infrastructure isn't ready.

Do Caos do OpenClaw ao Frontier da OpenAI: O Ajuste de Contas da Infraestrutura de Agentes

OpenClaw foi de sensação viral a pesadelo de segurança. Hoje, a OpenAI lançou o Frontier. Dois lados da mesma história: agentes autônomos estão chegando ao mainstream, e a infraestrutura não está pronta.

The 100th Tool Call Problem: Why Most AI Agents Fail in Production

Demos run for minutes. Production runs for hours. Here's what breaks, and the durability patterns that actually work when your agent needs to make 200+ tool calls without losing its mind.

O Problema da 100ª Chamada: Por Que a Maioria dos Agentes de IA Falha em Produção

Demos rodam por minutos. Produção roda por horas. Aqui está o que quebra, e os padrões de durabilidade que realmente funcionam quando seu agente precisa fazer mais de 200 chamadas de ferramenta sem perder o rumo.

Specs Are Not Free: Why AI Won't Replace Programming (It Will Transform It)

A response to the viral "spec as code" thesis. We don't stop programming—we change what programming means. And that change requires more engineering fundamentals, not fewer.

Specs Não São de Graça: Por Que a IA Não Vai Substituir a Programação (Mas Vai Transformá-la)

Uma resposta à tese viral "spec as code". Não paramos de programar - mudamos o que programar significa. E essa mudança exige mais fundamentos de engenharia, não menos.

The Agent Harness: Why 2026 is About Infrastructure, Not Intelligence

Intelligence without infrastructure is just a demo. Here's why the Agent Harness thesis matters and what I've learned building autonomous agents that actually work.

O Agent Harness: Por Que 2026 é Sobre Infraestrutura, Não Inteligência

Inteligência sem infraestrutura é apenas uma demo. Veja por que a tese do Agent Harness importa e o que aprendi construindo agentes autônomos que realmente funcionam.

2025: The Year AI Agents Got Real

A practitioner's look back at what happened in AI agents in 2025—from DeepSeek shaking the markets to GPT-5, Gemini 3, and the rise of agentic AI.

2025: O Ano em que os Agentes de IA Ficaram Reais

Um olhar de quem constrói sobre o que aconteceu com agentes de IA em 2025—do DeepSeek abalando os mercados ao GPT-5, Gemini 3 e a ascensão da IA agêntica.

The Memory Problem in AI Agents

Why memory architecture is the hidden bottleneck in AI agent systems, and the patterns that actually work in production.

O Problema da Memória em Agentes de IA

Por que a arquitetura de memória é o gargalo oculto em sistemas de agentes de IA, e os padrões que realmente funcionam em produção.

Where's the shovelware? Right here. Why AI coding works (if you know how to use it)

Real data showing 1082% productivity gains with AI coding tools. GitHub metrics, workflow diagrams, and evidence-based response to AI coding skeptics with 146,000 lines shipped in 4 months.

Software produzido em massa? Temos. Por que programar com IA funciona (se você souber usar)

Dados reais mostrando ganhos de produtividade de 1082% com ferramentas de codificação IA. Métricas do GitHub, diagramas de fluxo e resposta baseada em evidências com 146.000 linhas entregues em 4 meses.

Why I stopped estimating: a data-driven case against software predictions

A data-driven analysis of why software estimation doesn't work and what to do instead. Based on research, industry reports, and real-world experience with the no-estimates movement.

Por que parei de estimar: um argumento baseado em dados contra previsões de software

Uma análise baseada em dados sobre por que estimativas de software não funcionam e o que fazer no lugar. Baseado em pesquisas, relatórios da indústria e experiência real com o movimento no-estimates.

How to strike the right balance between not rushing to show value and exceeding expectations in early-stage startups

Explore the significance of balancing value and a co-founder mindset in early-stage startups. Learn how to show value while respecting team dynamics and collaborating towards common goals.

Como encontrar o equilíbrio entre não se apressar para mostrar serviço e superar expectativas em startups early stage

Descubra como encontrar o equilíbrio ideal em startups de estágio inicial, equilibrando a demonstração de valor, superação de expectativas e respeito aos processos existentes.

Creating an efficient and healthy interview process for engineers: insights from a decade of leading engineering teams

Create an efficient interview process for engineers. Learn practical tips for assessing skills, fit, and diversity, attracting top talent, and mitigating biases.

Criando um processo de entrevista eficiente e saudável para engenheiros: dicas práticas e sugestões

Descubra como criar um processo de entrevista eficiente e saudável para engenheiros. Saiba como avaliar habilidades técnicas e ajuste cultural, remover vieses e promover a diversidade. Exemplo prático e dicas para criar um ambiente inclusivo.

The importance of zero trust architecture for enterprises with remote or hybrid work models

Learn about Zero Trust Architecture for remote and hybrid work. This approach to cybersecurity provides stronger security and lower risk for businesses.

A importância da arquitetura zero trust para empresas com modelos de trabalho remoto ou híbrido

Saiba mais sobre a Arquitetura Zero Trust para trabalho remoto e híbrido. Essa abordagem de cibersegurança oferece maior segurança e menor risco para empresas.

The world is not that bad! Book review: Factfulness

Explore ‘Factfulness’ by Hans Rosling and learn how the world is improving. Discover how our instincts can distort reality and read about global progress.

O mundo não é tão ruim assim! Review de livro: Factfulness

Explore ‘Factfulness’ by Hans Rosling and learn how the world is improving. Discover how our instincts can distort reality and read about global progress.

A brief look at hey.com stack

An exploration of the technology stack and infrastructure behind Basecamp's Hey email service, including Ruby on Rails, Kubernetes, and their development workflow.

Uma olhada rápida na stack do hey.com

Uma exploração da stack tecnológica e infraestrutura por trás do serviço de email Hey da Basecamp, incluindo Ruby on Rails, Kubernetes e seu fluxo de desenvolvimento.

Avoiding burnout as a software engineer

Learn how to recognize, prevent and fight burnout as a software engineer, with practical tips for both individuals and managers to create healthier work environments.

Evitando burnout como engenheiro de software

Aprenda como reconhecer, prevenir e combater o burnout como engenheiro de software, com dicas práticas tanto para indivíduos quanto para gestores criarem ambientes de trabalho mais saudáveis.

What diversity should not be about, especially in software teams

A thoughtful perspective on diversity in software teams, exploring what true diversity means and what practices should be avoided when promoting inclusivity.

Como não se deve promover diversidade, principalmente em times de software

Uma perspectiva reflexiva sobre diversidade em times de software, explorando o que realmente significa diversidade e quais práticas devem ser evitadas ao promover inclusão.

My first exit: what I've learned in 8 years leading meuingresso.com

The story of my first startup exit after 8 years building and leading meuingresso.com, including lessons learned about execution, trust, and the entrepreneurial journey.

Meu primeiro exit: o que aprendi nos 8 anos à frente do meuingresso.com

A história do meu primeiro exit após 8 anos construindo e liderando o meuingresso.com, incluindo lições aprendidas sobre execução, confiança e a jornada empreendedora.

How to develop reusable components with Babel and RollupJS

A comprehensive guide on creating and publishing reusable JavaScript components using Rollup.js and Babel, including the evolution of JavaScript modules.

Como desenvolver componentes reutilizáveis utilizando Babel and RollupJS

Um guia completo sobre criação e publicação de componentes JavaScript reutilizáveis usando Rollup.js e Babel, incluindo a evolução dos módulos JavaScript.

var, let or const?

Understanding the differences between var, let, and const in JavaScript ES6, including scope behaviors and when to use each variable declaration.

var, let ou const?

Entendendo as diferenças entre var, let e const no JavaScript ES6, incluindo comportamentos de escopo e quando usar cada declaração de variável.

My personal mission statement

My personal mission statement defining how I approach life, work, relationships, and personal growth, inspired by Stephen Covey's principles.

Declaração de missão pessoal

Minha declaração de missão pessoal definindo como abordo a vida, trabalho, relacionamentos e crescimento pessoal, inspirada pelos princípios de Stephen Covey.

Beber líquidos durante as refeições atrapalha a digestão. Mito ou verdade?

Uma análise científica baseada em evidências sobre os efeitos da ingestão de líquidos durante as refeições na digestão e saúde.