Software Engineer (Model Inference)
Quick Summary
About us We're building a new category of interactive entertainment powered by AI characters, image, video, and real-time experiences. We’re one of the most-visited AI products globally,
We're building a new category of interactive entertainment powered by AI characters, image, video, and real-time experiences. We’re one of the most-visited AI products globally, serving millions of users every day.
We work in-person in Sydney, Australia, and hire globally.
User-first: We build what people want. We invest time to understand our users and focus on adding value instead of extracting value.
High agency, high ownership: We own outcomes end-to-end. When something goes wrong, we take responsibility and fix it.
Urgency: We prioritize ruthlessly, increase leverage, and move at an exceptional pace.
Responsibilities
~1 min read- →
You'll own our model inference stack end-to-end — serving our in-house and open-source models to tens of millions of users at low latency and high throughput, and squeezing every bit of performance out of our GPU fleet.
Ship a high-throughput inference server for our in-house models, serving millions of generations per day at <200ms latency.
Squeeze more out of every GPU with batching, quantization, and custom CUDA kernels.
Test and productionize LoRAs and our in-house models to serve 10s of millions of users.
5+ years of experience building software at scale, with a focus on ML inference or GPU-accelerated systems.
Deep familiarity with GPU inference: batching, quantization, and serving frameworks like vLLM, TensorRT, or Triton.
The ability to get shit done end-to-end. From understanding users → proposing an idea → implementation → iteration.
Hunger to win. This is not going to be easy.
What We Offer
~1 min readWe raise the ceiling for exceptional performers, not the floor. As your impact grows, your compensation should too.
The listed range is cash + equity, with superannuation on top.
We review compensation every 6 months.
Bonuses reward past impact, and raises reflect your new level.
Intro call: 15 minutes to align on the role and company.
Founder interview: go deep on what you’ve built, how you think, and how you work.
Technical interview: 1 hour of system design and technical discussion. No live coding.
Paid work trial: spend 3 days with us in Sydney working on a real problem. We cover travel and accommodation, if needed.
Offer: if it’s a strong fit, we move quickly.
Location & Eligibility
Listing Details
- Posted
- July 28, 2026
- First seen
- July 28, 2026
- Last seen
- September 10, 2026
Posting Health
- Days active
- 0
- Repost count
- 0
- Trust Level
- 63%
- Scored at
- July 28, 2026
Signal breakdown
Please let coreflow know you found this job on Jobera.
3 other jobs at coreflow
View all →Explore open roles at coreflow.
Similar Software Engineer jobs
View all →Browse Similar Jobs
Stay ahead of the market
Get the latest job openings, salary trends, and hiring insights delivered to your inbox every week.
No spam. Unsubscribe at any time.