AI Engineer
Company: Cozi
Location: Indiana (Remote)
Type: Full-time
Remote: Yes
Posted: 2026-08-07
About this role
At In Tandem, we build technology that helps families manage everyday routines and navigate life’s biggest transitions. Through our four brands—OurFamilyWizard, Cozi, FamilyWall, and Custody Navigator—we help families stay organized, communicate well, and foster healthy childhoods.
We believe technology should strengthen relationships and make daily coordination less complicated. Everything we create is designed to lighten the mental load, reduce conflict, and support families through big and small moments.
If you want your work to make a real difference in the daily lives of parents and kids, In Tandem is the place where your impact will truly matter.
As our
AI Engineer
, you'll keep the AI infrastructure our products and teams run on fast, efficient, and reliable, and you'll build with it. You'll run and optimize our self-hosted inference stack on our own GPU hardware, build the internal platform our employees work through, and ship user-facing agents inside the apps. Your work spans OurFamilyWizard, Cozi, and FamilyWall, and the platforms that power how we build.
This is a hands-on technical role at its core: you own the technical side of running our models on our own hardware. But it's not siloed, and we don't want it to be. We're looking for someone who also wants to pick up app-layer work and ship product-facing features, and does both well.
*Run and optimize our self-hosted inference stack*
- Run the inference serving layer on our own GPU hardware: choose and tune the serving stack (vLLM, SGLang, TensorRT-LLM) for high throughput and low latency.
- Optimize aggressively: tensor parallelism, quantization (FP8, AWQ, GPTQ), KV-cache and prefix caching, continuous batching, speculative decoding, concurrency tuning.
- Serve multiple models and features off shared hardware: multi-LoRA, routing, and request scheduling that balances internal workloads against latency-sensitive product traffic.
*Keep our AI fast, efficient, and obse...