Robotics · AI/ML · Software

Jaiveer Bassi

I build and audit intelligent systems, from robot data pipelines and fault injection rigs to reproducible machine learning research.

I am an AI Robot Operator at Physical Intelligence and Founder & Executive Director of House of Seva. I hold an M.S. in Software Engineering and a B.S. in Computer Science.

Open to full time robotics software, AI/ML, and software engineering opportunities.

Portrait of Jaiveer Bassi

Current focus

Research

Reproducible work with the claim, evidence, controls, and limitations kept together.

Published and locally reproduced HumanEval scores for five code models
Technical reportCode models · Local inference

Do Published HumanEval Rankings Survive Local Deployment? A Five-Model Reproduction on a Consumer GPU

Reproduces a published five model ordering under one pinned RTX 4070 evaluation stack. Four models retained their relative order, while StarCoder2 3B reversed with Qwen2.5 Coder 0.5B, reducing Kendall’s tau b to 0.8. Version 1.0 is a technical report and has not been submitted to a venue.

FP32 and INT8 latency distributions with backend specific speedups
Archived preprintEdge AI · Host CPU inference

When Does INT8 Actually Accelerate Host-CPU Inference? A Reproducible Audit of MobileNetV2 Conversion and Packaging

Finds that INT8 is smaller but not intrinsically faster or accurate. Per channel INT8 accelerated one thread XNNPACK inference by 1.27 times, yet ran 2.24 times slower without the default delegate. Version 1.2.0 is archived on Zenodo and has not been submitted for peer review.

Directional PGD transfer rates across budgets and model pairs
Under review at TMLRAdversarial ML · CIFAR 10

Conditioning and Directionality in Adversarial Transfer on CIFAR-10

A denominator aware study across ResNet 18, VGG16, and MobileNetV2 with three training seeds, full test evaluation, sensitivity analyses, and prediction level artifacts. Submitted to Transactions on Machine Learning Research through OpenReview.

Nearly identical MRI scans found across benchmark test and training folders
Preprint · Not journal submittedMedical imaging · Benchmark audit

Complete Patient-Level Leakage Among Traceable Images in a Widely Used Brain Tumor MRI Benchmark

Finds that every traceable tumor test image belongs to a patient also present in training and that 216 test images are pixel identical to training images. The release provides recovered patient identifiers and patient disjoint folds. The manuscript has not been submitted to a journal.

Closed loop success and open loop error across five demonstration failure modes
Submitted to CoRL 2026 LfCRobot learning · Data quality

Not All Bad Demonstrations Are Equally Bad: Quantifying How Demonstration Failure Modes Degrade Closed-Loop Policy Performance

A controlled study of 660 policy fits across five failure modes, two policy families, ten seeds, and transition matched controls. It shows why open loop prediction error can misrepresent closed loop policy damage. Submitted to the CoRL 2026 Learning from Corrections and Interventions workshop.

Trajectory smoothness metrics under episode length controls
Three workshop submissionsRobot learning · Demonstration curation

Smoothness Ranks Skill, Not Success: An Audit of Published Trajectory-Smoothness Curation Metrics Under Episode-Length Controls

Audits SAL and TED on robomimic and finds that the metrics primarily rank operator skill rather than task success. Venue specific versions are under review at CoRL WEBP, NeurIPS RoboPAD, and CoRL Oops, I Erred.

Credentials

Certifications

Selected technical and professional credentials.

Writing

Articles

CAN bus fault diagnostics rig

Simulating Real Robot Failures on a Breadboard

A three node CAN bus built to induce and diagnose failures that software logs alone cannot explain.

Read on Medium →
Low cost leader follower teleoperation rig

A $30 Teleoperation Rig

Rebuilding the core idea behind GELLO style robot demonstration collection with direct joint space control.

Read on Medium →

Timeline

Experience & education

Physical Intelligence

AI Robot Operator

Physical Intelligence · May 2026 to present

House of Seva

Founder & Executive Director

House of Seva · June 2026 to present

Handshake AI

LLM Response Evaluator

Handshake AI · April to June 2026

Kaiser Permanente

Software Developer

Kaiser Permanente · October 2023 to June 2024

M.S., Software Engineering

Grand Canyon University · 2024 to 2025

University of Silicon Valley

B.S., Computer Science

University of Silicon Valley · 2021 to 2023

Beyond work

Hobbies

Reading and philosophy

I read philosophy, theology, science, and the classics. Add me on Goodreads to see what I am reading, and follow my writing on Medium.

Building keyboards

I enjoy building custom keyboards and plan to release projects and articles about them whenever possible. Stay tuned.

3D printing

I design, print, and refine physical pieces through Castline Studio, combining digital fabrication with hands on experimentation.

Build, break, measure

Selected projects

Robotics infrastructure, fault characterization, ML tooling, and production software. More on my GitHub.

Robot telemetry pipeline dashboard

Kubernetes · Redis Streams · PostgreSQL

Fault Tolerant Robot Telemetry Pipeline

A multi stage pipeline designed for deliberate pod and network failures, with durable acknowledgements, replay safe writes, health probes, observability, and a chaos harness.

GitHub →
UR5e RTDE harness terminal demonstration

URSim · RTDE · Fault injection

UR5e RTDE Harness

Simulator backed control and synchronized telemetry with reproducible tests for connection failures, stale handles, hangs, and native crashes.

GitHub →
Habitua iOS application screens

SwiftUI · AVAudioEngine · DSP

Habitua

A dependency free tinnitus habituation app with real time generated sound enrichment, personalized notch filtering, guided exercises, and progress tracking.

GitHub →
Fleet Triage robot health dashboard

Python · Streamlit · Diagnostics

Fleet Triage

Robot fleet log analysis with fault classification, recurrence detection, anomaly detection, rule based root cause analysis, and a live health dashboard.

GitHub →
Robot grasp annotation workflow

Computer vision · Labelbox · COCO

Grasp Annotation Pipeline

A data operations workflow for robot arm grasp annotations with bounding boxes, keypoints, failure labels, and automated COCO JSON export.

GitHub →
AI Model Packager command line demonstration

Python · Docker · MLOps

AI Model Packager

A library and CLI for packaging machine learning models into portable, optimized Docker inference images.

GitHub →
Chess reinforcement learning interface

C++ · PyTorch · Self play

Chess RL Engine

An AlphaZero inspired system combining a C++ bitboard engine, PyTorch neural network, self play training, and a PyGame interface.

GitHub →
Live CareerTuner resume analysis homepage

OpenAI · React · Node.js · Supabase

CareerTuner

An AI resume analyzer with job specific feedback, ATS scoring, skill gap analysis, and PDF export.

Live site →
Live MedStract biomedical literature analysis homepage

Python · PubMed · NLP

MedStract

A biomedical literature tool for searching PubMed, rewriting summaries for different audiences, exporting citations, and generating PDFs.

Live site →
Castline Studio order interface

Next.js · TypeScript · WebAssembly

Castline Studio Order System

An e commerce workflow with client side STL analysis, Etsy and shipping integrations, automated email, and deployment on Vercel.

Live site →