Neo Mohsenvand — AI/ML Engineer and Product Designer

I'm an AI/ML engineer and product designer, and Chief AI Architect for Global Finance at Caterpillar, where I advance AI adoption through original products, shared infrastructure, and engineering education. I build enterprise AI agents and the software around them: document workflows, evaluation systems, voice interfaces, and tools that bring people into decisions. I also make local AI tools and interactive educational applications for exploring agent workflows, training models, and understanding retrieval. Before Caterpillar, I led research at BrainCo and worked on machine learning for biosignals at Apple. I did my PhD at MIT, in Human–AI Interaction at the Media Lab, and stayed on there as a postdoctoral researcher; my work explored self-supervised learning, brain-computer interfaces, and tools for human memory and attention. Before MIT, I earned a master's in Mathematical Modelling and Scientific Computing at Oxford. I'm drawn to the intersection of software, design, and intelligence, both artificial and biological.

Experience

  1. Chief AI Architect, Global FinanceJanuary 2026 – present · Irving, TX

    • Tempo: Built an enterprise investigation platform and designed its meeting-like interview interface, combining stakeholder discovery, scheduling and OpenAI Realtime voice interviews with structured forms. Available to hundreds of internal users for idea intake, requirements gathering and conflict resolution.
    • Aperture: Built a new agentic image-generation harness for slide decks and visual assets, using terse prompts and hundreds of references to reduce reliance on third-party subscriptions.
    • Touchless Invoice Processing agent: Co-developed end-to-end orchestration from email ingestion and scanned-invoice OCR through exception handling and Snowflake writes, with a detailed multi-level evaluation framework. Reduced invoices needing human review from 100% to 2%.
    • Teaching: Created 10+ workshop apps, including agentic RAG, talking dashboards and a voice-only data-science interface I designed. Teach engineers and senior leaders in the workshop's fifth iteration.
    • Developing shared infrastructure for AI application prototyping and deployment across Global Finance.

    Senior Principal Data Scientist, Global FinanceMay 2025 – January 2026 · Irving, TX

    • Penman: Built and deployed an enterprise financial-analysis system using SvelteKit and Azure to analyze earnings-call transcripts and regulatory filings for finance leaders.
    • Redesigned earnings-call transcript proofreading in Penman to help finance teams identify disclosure risks.
    • Built MLflow / DeepEval evaluation infrastructure and coordinated organization-wide benchmark development.
    • Built an AI-idea prioritization agent with RICE-derived scoring, internal impact metrics and human review.
    • Designed adversarial LLM analysis of draft financial disclosures.

    Senior Engineering Fellow, Artificial IntelligenceApril 2024 – May 2025 · Illinois

    • Founding technical hire of CAT Tech AI; led four engineers delivering 12+ applications for 8,000+ technical users through engineering-analysis automation.
    • Created Magic Table for million-row spreadsheet/database workflows, combining formulas and iterative local-LLM prompts to reduce cloud compute costs; developed versions with Ollama.
    • Built a ModernBERT-based embedding model for Caterpillar products and parts; orchestrated data collection, developed a synthetic training-pair generation app, and trained on an on-premises H200 cluster.
    • Led distributed training, data curation, SFT/PEFT, PPO/DPO alignment, and automated and human evaluation.
    • Established an AI curriculum and reading group across 10+ teams, training 300+ engineers.
    • Led 11 researchers developing wearable AI and brain-computer interfaces; named inventor on five provisional patent applications.
    • Led development of a conditional-GAN system reconstructing breathing patterns from 25 Hz wearable PPG. The published study reports a one-second breathing-curve feedback delay and respiratory-rate mean absolute error of 1.47 breaths/min across ten test subjects.
    • Led soft-speech interface research comparing contact and MEMS microphones. In an eleven-person study, contact microphones reduced soft-speech word error rate from 99–100% to 35% in the tested music/noise conditions.
    • Researched multimodal memory systems and physiological data, building on doctoral work in continuous-capture memory interfaces and self-supervised EEG representations.
    • Developed parameter-efficient transformers for biosignal event detection and continuous time-series processing.
    • Built methods to inspect attention heads and activation patterns in continuous time-series models.
    • Built a browser-based WebGL visualization engine for dynamic social-network graphs.
    • Developed a Python library for retrieving, composing, and visualizing Allen Brain Atlas connectivity data in 3D.

Education

    • Completed in 2021; listed in MIT’s February 2022 degree record. Research in self-supervised learning and human–AI interaction.
    • Thesis: Classifying and Displaying Brain-Waves through Self-Supervised Learning. Advisor: Pattie Maes. Co-advisors: Tomaso Poggio and Ed Boyden.
    • First author of SeqCLR (ML4H NeurIPS Workshop, 2020), adapting contrastive learning to EEG through channel recombination and signal augmentations; evaluated emotion recognition, sleep staging, and abnormal-EEG detection across three datasets.
    • Doctoral work also included continuous-capture memory interfaces and tools for human attention.
    • Sparse-Signal Recovery and Compressive Sensing for Optical Functional Brain Imaging.
    • Advisor: Jared Tanner. Co-advisor: Mason Porter.
  1. Tehran Polytechnic2011 BSc Biomedical Engineering · BSc Electrical EngineeringPresident, IEEE Student Branch

Honours

  • NTT Data Fellowship — MIT, for Rhizome, a smart memory book
  • First place and the Koch AI Prize — MIT Grand Hack, for Memoroom, VR for Alzheimer’s
  • Bill Mitchell Design Award — MIT Media Lab
  • Khwarizmi Prize — first in mathematics and physics, and third the year before
  • Semifinalist — national mathematics and physics olympiads

Skills

Programming
Python, TypeScript, MATLAB, Mathematica, GLSL, WGSL
Applications & UI
SvelteKit, React, Three.js / WebGL
Design & UX
Adobe Creative Cloud, UX study design, Design Systems
Cloud & data
Azure, AWS, Docker, progressive delivery, Snowflake, multimodal data pipelines
Data science
Time-series analysis, biosignal processing, feature extraction, scientific visualization
Machine learning
PyTorch, JAX, transformers, self-supervised learning, mechanistic interpretability
Training & alignment
FSDP, DeepSpeed, Ray; SFT, PEFT, LoRA, PPO / DPO, quantization
Inference & serving
Ollama, Transformers.js, WebGPU, ONNX / WASM, vLLM
Realtime AI
OpenAI Realtime, ElevenLabs, Deepgram; low-latency multithreaded harnesses
Agentic engineering
LangGraph, agent harnesses, RAG, tool orchestration, human-in-the-loop workflows
Evaluation
MLflow, DeepEval, benchmark design, automated and human evaluation
Leading & teaching
AI system architecture · cross-functional R&D · curriculum design and workshops · mentoring · technical recruiting
Languages
English · Turkish · Persian · Mandarin

Work

32 of 32

Tissue

An interpretability lab: train a transformer in your browser, then probe its neurons.

AIResearchWebGPU

Pattern

Machine learning for absolute beginners, from fitting a line to running an agent.

EducationalAIWebGPU

Before the Move

World models from first principles — train one in the page, or ask the book aloud.

EducationalAIWebGPU

harnessXray

See how an AI agent really works: prompts, tools, memory, and token cost.

EducationalAIResearch

Voicebook

Talk with your documents: AI-controlled scrolling and highlighting, local inference, and optional cloud models.

AIWebGPUTools

LangX

An interactive AI engineering course: 29 lessons on the LangChain stack.

EducationalAI

jaxverse

An interactive book on deep learning where every model trains live on your own GPU.

EducationalAIWebGPU

Embedding Playground

Five labs for text-embedding models: compare, trace, search, classify, cluster.

EducationalAIWebGPU

gradientlab.ai

Watch optimizers navigate loss landscapes in real time, with D3-powered animations.

EducationalAIResearch

VLMOCR

Multi-region image-to-text with vision language models. Prompt and stream.

AIResearch

jax-js skill

An agent skill for training real neural networks in the browser with jax-js.

AIWebGPU

General Relativity

From a cart and a clock to curved spacetime — change the experiment, or ask the page aloud.

EducationalAIWebGL

MIDI Lab

Thirty-one lessons wired to a live MIDI engine — press a key and the bytes appear, decoded.

EducationalTools

TerminalVibes

A visual guide to the shell for the AI era. Read, verify, and run in a browser playground.

Educational

Protocol Lab

An interactive atlas of 46 network protocols, from TCP to QUIC.

Educational

GitVibes

A visual, interactive guide to Git for AI-assisted coders, minus the jargon.

Educational

SoftMax Explainer

The softmax function made interactive: temperature, logits, probability.

Educational

Neocalculus

An infinitesimal-first calculus book. Learn by manipulation, not memory.

Educational

RDB

Relational databases from first principles — with live SQL playgrounds in every section.

Educational

UI Atlas

A component reference for vibe coders: browse, compare, copy in one place.

Educational

Swarm

Thousands of agents forming emergent flocks, GPU-accelerated in your browser.

ResearchArtArtificial LifeWebGPU

Moiré

A WebGPU studio for interference fields, where every layer is a scalar field rather than a drawing.

ArtWebGPU

Games of Life

WebGPU-powered cellular automaton engine with a live rule editor, presets, and brush tools.

ArtArtificial LifeWebGPU

Algocell

Artificial life emerging from Z80 machine code bytes — evolution in your browser.

Artificial LifeWebGPU

Zilion

Thousands of Z80 CPUs in a single WebGPU dispatch, an emulator built for artificial life.

WebGPU

Boids

Craig Reynolds' flocking algorithm in WebGL, every rule a live control.

ArtArtificial LifeWebGL

Vibe Coding

A first-person flight through volumetric clouds and ocean waves, raymarched in real time.

ArtWebGPUWebGL

Moiré Fields

The theory behind Moiré: interference as a scalar field, and where the fringes fall.

ResearchArt

Goldbach

One FFT counts every Goldbach partition — and the Riemann zeros turn up in the residue.

Research

MacMac

A game about sampling probability distributions with the fewest clicks.

GameResearch

Vid2GIF

A native macOS video-to-GIF converter that previews the real encoded output as you tune it.

Tools

TalkOver

Record a browser tab with mic and webcam overlay, then export to video or GIF. All local.

Tools

Earlier work

inSight: Deep Neurofeedback Brain decoding using BigGAN and MusicVAE to generate naturalistic neurofeedback stimuli.
Affective Memory Summarization Using affective signals to condense 16 hours of body-cam footage into a 15-minute daily recap.
Large-Scale EEG Biometrics Self-supervised EEG-based user identification at scale, surpassing prior 157-subject limits.
Physiophone Sonification of electrophysiological signals — turning EEG, ECG, and EMG into sound.
Flower: EEG Visualization Open-source tool for in-depth visual analysis of multi-channel time-domain neural recordings.
SkipNorm A flexible deep learning building block that combines skip connections and normalization in one layer.
Q: Conversational Data Labeling An intelligent interface that builds a personal knowledge graph through natural conversation.
DataVRse VR visualization of the Twitter social graph during the 2016 U.S. Presidential Election.
Rhizome A tool for navigating personal memories to support people living with dementia.
The Electome AI-driven mapping of the 2016 election public sphere — where machine intelligence meets political journalism.
SeqCLR: Self-Supervised Features Contrastive learning of representations for EEG and other time-series data, without labels.
VR Maze in Zero Gravity Testing how spatial memory holds up on a parabolic flight, once the sense of up is gone.
The Foodome A knowledge graph of food, linking how we talk and learn about it to what it actually contains.
The Electome (Talk) A talk on the Electome from the Laboratory for Social Machines, built with Twitter and Knight Foundation.
Sole2Soul An interface with the soul of a city. First place and the Bill Mitchell Design Award at Make Me++.