I build computer vision and multimodal AI products, and the inference infrastructure that puts them into production.

Faces in moving cars. Tissue on pathology slides. Then the GPU serving layer and agent runtime underneath models like them. Thirteen years of teaching machines to read the physical world, and of finding out what it costs to keep that running in front of real customers.

The lesson repeats in every domain. Capability is the easy part. Unit economics, integration, regulation, and trust are the work.

Work

Writing

I write Out of Frame at writing.abdomahmoud.com, about multimodal AI, perception systems, and the infrastructure that runs them in production. You can subscribe there.

Bylines

Get in touch

If you're building computer vision or multimodal AI products, or the infrastructure that runs them at scale, I'd like to hear from you. Find me on LinkedIn.