
Software Engineer, Applied AI
New York, NYFull-time$175–290K/yrPosted 5mo agoStill listed 4w ago
Most applications go out cold — see where you stand first. No sign-up to start.
Watch jobs like this. New roles like this one near New York, NY, by email.
Don't just apply. Show up ready.
Olive works from this exact posting.
At a glance
Olive lists jobs from US employers, including remote roles you can work from the United States.
Job overview
Auctor is hiring a Software Engineer, Applied AI. Auctor is building the AI layer for professional services and software implementation, creating core systems that power agents in production. The role focuses on designing, building, and improving retrieval, document understanding, tool use, memory, and orchestration within an empirical research environment.
Key focus areas include Build and improve core systems behind agents across retrieval, tool use, document understanding, memory, and orchestration, Design evals and experiments to understand agent quality in production, and Turn traces, failures, and user behavior into concrete product and architecture decisions.
Successful candidates bring Strong Engineering Fundamentals and Ability To Ship Production Systems. Important skills include Python, Building LLM-Powered Products, Building Agent Systems, Applied AI Systems, Empirical Mindset, and Systems Taste. Preferred (not required): Retrieval Systems, Search Systems, Ranking Systems, and Designing Evals.
Skills & qualifications
Skills
Qualifications
Benefits
Full job description
WHY AUCTOR
Auctor is building the AI layer for professional services and software implementation. Think of us as the brain behind the best solution engineers, forward-deployed engineers, and onboarding teams—automating the documentation, the discovery, and the decision-making that powers $400B+ in services work. We're going after one of the biggest software categories of the decade.
ROLE OVERVIEW
As a Software Engineer, Applied AI at Auctor, you will design, build, and improve the core systems behind our agents in production.
This role sits at the boundary of engineering and empirical research. You will work across retrieval, document understanding, tool use, context management, prompting, and orchestration. Some weeks you will be shipping new capabilities. Some weeks you will be mining production traces, designing evals, and figuring out which part of the system is actually failing.
We are not looking for someone to glue an API onto a product and call it AI. We are looking for someone who wants to build real agent systems, understand how they behave in the wild, and use that understanding to make bold product and architecture decisions.
This role is based in New York, NY, in person 5 days per week.
WHAT YOU'LL DO
-
Build and improve the core systems behind our agents across retrieval, tool use, document understanding, memory, and orchestration
-
Design evals and experiments that help us understand agent quality in production
-
Turn traces, failures, and user behavior into concrete product and architecture decisions
-
Work closely with operations, GTM, and deployed teams to understand real workflows and where agents break down
-
Evaluate models, prompts, and system designs across real enterprise tasks
-
Own the loop from idea -> implementation -> measurement -> iteration
WHAT WE'RE LOOKING FOR
-
Strong engineering fundamentals and the ability to ship production systems
-
Fluency in Python
-
Experience building or working on LLM-powered products, agent systems, or adjacent applied AI systems
-
An empirical mindset — you reach for logs, traces, experiments, and real usage before guessing
-
Strong systems taste — you understand that retrieval, prompting, memory, tools, and UX interact
-
High ownership and comfort working in ambiguity
-
Strong opinions about what makes agent systems actually work
STRONG CANDIDATES MAY ALSO HAVE
-
Experience with retrieval, search, or ranking systems
-
Experience designing evals, benchmarks, or feedback loops for LLM systems
-
Experience building internal tools, workflow products, or operator-facing systems
-
Experience in startups or other high-ownership environments
EXAMPLE PROJECTS
This is a new field. We care much more about what you have built than whether your background fits a standard template.
Projects that would make us excited include:
-
Designing and shipping an agent harness that materially improved performance on a real task
-
Building an eval or benchmark that changed what your team decided to build next
-
Designing tool interfaces, memory systems, or retrieval systems for an LLM-powered product
-
Building a production workflow around language models that users actually depended on
-
Running a careful experiment on prompting, model routing, or orchestration and using it to drive a product decision
If you apply, we would love to see one thing you built with LLMs or agents. It does not need to be perfect or flashy. We mostly want to understand how you think, what you owned, what you learned, and what tradeoffs you made.
COMPENSATION
$175,000-$290,000 base salary, plus equity.
BENEFITS:
-
Equity with real upside – you're joining early, and your equity reflects that
-
Competitive, top-of-market salary
-
Medical, dental, and vision coverage
-
Monthly wellness stipend
-
Unlimited PTO – take the time you need
-
Daily meal stipend, plus dinner covered on the nights you're working late
-
Weekly happy hours and team outings because work is better with people you actually like
Similar jobs, posted recently
Open roles like this one, listed in the last 30 days.
Software Engineer, Applied AIMercor · New York, NY · $130–500K/yrPosted 2w agoPosted 2w agoApplied AI EngineerMaintainX · San Francisco, CA (Hybrid) · $152–271K/yrPosted 2w agoPosted 2w ago
Lead Applied AI EngineerLangChain · New York, NY · $160–200K/yrPosted 2w agoPosted 2w ago
Senior Software Developer, LiveDesign InfrastructureSchrödinger · New York, NY (Hybrid) · $155–270K/yrPosted 4 days agoPosted 4 days ago
Software EngineerAdobe · New York, NY · $139–258K/yrPosted 1w agoPosted 1w ago
You've read the whole posting — now see how you match it.