I’ve been making things for a long time. While the destination was always fantastic - little is as catharic as making something truly novel exist - the journey is oft as important. My skills and experience that have powered my career …
I’ve long been a fan of task runners within projects - I find they’re practically required once you have a project or group operating beyond a certain level of complexity. There are just too many one liner terminal commands to …
tl;dr I talk about the ideas behind and efforts expended to build my AI framework, arkaine. What worked? What didn’t? Does it have a future?
arkaine, briefly I decided a few months ago to create an AI framework for making agents. I …
If I had the budget/compute power and time to research anything at the moment, this is what I would try researching. It’ll make sense in a bit, I swear.
The Problem Training robotic policies to perform complex tasks is incredibly …
tldr; A paper caught my eye, proposing a way to treat individual agent calls in multi-turn agent training as individual steps. The key contribution - an easy, intuitive method of assigning credit to each step, for quicker training.
[Paper] …
Recently, I gave a talk on several of DeepSeek’s innovations, which were as extensive as they were complicated. The particular clever discovery that best captured my imagination was their development of GRM and SPCT.
Most AI models — …
tldr I recently gave a talk at the SDx paper club which I run on the paper DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning. I wanted to take a moment to blog out some of the talk, specifically how their …
Your browser does not support the video tag. Google Deepmind recently released Diffusion Models Are Real-Time Game Engines [Site] | [Paper], a fascinating paper wherein a modified Stable Diffusion model acts as the game engine for the …
tl;dr Google DeepMind released a paper claiming that, without search, a transformer architecture can be utilized to achieve grandmaster level play in chess. But what does that mean? I’m giving a paper club talk on the subject, so I …
tl;dr A recent paper studied large language model’s (LLM) reactions to stimuli in a manner similar to neuroscience, revealing an enticing tool for controlling and understanding LLMs. I write here about the paper and perform some …
tl;dr I got nerd sniped by a problem and wrote a solver for it. You can find it here.
The Problem Apparently it’s impossible to determine if someone is capable of coding without asking them arbitrary puzzles that are in no way related …
tl;dr For my Master’s capstone project, I demonstrated that not only can LLMs be used to drive robotic task planning for complex tasks given a set of natural language objectives, but also demonstrate some contextual understanding that …
Melvin Conway has a law that I’ve heard thrown about throughout my career building applications and backend systems. Verbatim it states:
Organizations which design systems are constrained to produce designs which are copies of the …
tldr ROS/ROS2 environments are notoriously annoying to get into a repeatable, isolated dev environment. I write about my initial look into this problem, and some proposed solutions I tried. I present my imperfect solution using both Docker …
tldr I write about some of the more interesting works that shaped my understanding of applying LLMs for AI agents and robotic applications.
Introduction What is this LLMs as a fad - a caveat Are LLMs actually going to be useful for …