My Neuro-Symbolic Meta-Policies paper was accepted at ISWC 2026
Why I want agents to choose their memory rules, not hide them
AI researcher and engineer.
Why I want agents to choose their memory rules, not hide them
Why I want a robot to remember how a team worked before
Why I care about learning which symbolic facts should persist
Why I wanted a lightweight, local-first way to use ArcadeDB from Python
Formal basics, real systems, and why industry favors property graphs while RDF remains important
Understanding the evolution from basic policy gradients to modern LLM fine-tuning algorithms
Understanding the evolution of generative models through practical implementations
Balancing n-gram models and neural networks for next-token prediction
Bridging Reinforcement Learning and Supervised Learning in LLMs
A Deep Dive into MLE, Loss Functions, and Beyond
Offline RL, Behavior Cloning, and The Magic of Sequence Modeling
Tracing the Layers from Machine Code to Natural Language Interfaces
Why Frequentist Significance Testing Falls Short
A natural language text can be seen as a knowledge graph
The reason why reinforcement learning is such a hard problem
4bit quantization is amazing
Exploring the challenges and strategies for effective edge classification with graph neural networks (GNNs)
Can machines think like us?
This really can do a lot of things, although it's still biased.
My boring old website had to be gone.
Unlocking the power of open-source language models
A simple room classifier made using EfficientNet and PyTorch Lightning
Revolutionizing NLP with a collaborative hub for crafting and sharing prompts
We won the NeurIPS 2021 competition
Achieving superior zero-shot generalization through explicit multitask prompting
Achieving state-of-the-rrt emotion recognition in conversations with a simple RoBERTa-based approach
A simple age-gender classification model using Arcface embeddings and MLPs