Simular logo

Software Engineer, CUA Control

Simular

SingaporeFull-time$100–205K/yrPosted 4mo agoChecked 2w ago

Most applications go out cold — see where you stand first. No sign-up to start.

At a glance

Compensation
$100–205K/yr
Location
Singapore
Schedule
Full-time
Work Authorization
Not specified

Job overview

Simular is hiring a Software Engineer, CUA Control. Simular seeks an engineer to develop the control layer that converts AI model intent into precise, reliable actions on computers, handling mouse and keyboard input, window management, UI detection, and error recovery across macOS, Windows, and Linux.

Key focus areas include Work on low-level computer control stack including mouse and keyboard injection, Implement UI element detection using accessibility APIs and visual grounding, and Build abstraction layer enabling cross‑OS and application type agent operation.

Preferred (not required): CGEvent, SendInput, xdotool, and AXUIElement.

Skills & qualifications

RequiredNice to have

Skills

CGEventSendInputXdotoolAXUIElementUI AutomationAt-SPIReliability EngineeringPlaywrightAppiumPyAutoGUIHammerspoonScreen Reader InternalsRDPVNCGame AutomationLLM Agent Tool-Use SystemsiOS UIAutomationXCTestAndroid UIAutomatorAndroid Accessibility

Qualifications

Built OS-Level Input AutomationUnderstands Accessibility FrameworksDealt With Flaky Element SelectorsMobile Device Automation Experience

Full job description

Where multiple locations are listed for this role, the position may be based in any of those locations, with priority determined according to the order of listing.

We're looking for an engineer to work on the control layer - the system that translates an AI model's intent into precise, reliable actions on a real computer. This means mouse movements, keyboard input, window management, UI element detection, and error recovery across macOS, Windows, and Linux.

What you'll do

  • Work on the low-level computer control stack: mouse/keyboard injection, screen capture, coordinate mapping, input simulation

  • Implement UI element detection using accessibility APIs (AXUIElement, UI Automation), DOM/a11y trees, and visual grounding

  • Help build the abstraction layer that lets our agent operate across OS platforms and application types

  • Tackle reliability problems: element targeting under UI changes, window occlusion, resolution scaling, cross-app focus management

  • Contribute to feedback loops: how does the agent know its action worked? How does it recover when something unexpected happens?

  • Work closely with the model and planning team on the interface between intent and execution

You might be a fit if

  • You've built OS-level input automation (CGEvent, SendInput, xdotool, or similar)

  • You understand accessibility frameworks - AXUIElement on macOS, UI Automation on Windows, AT-SPI on Linux

  • You've dealt with flaky element selectors, timing issues, resolution-dependent coordinates

  • You think carefully about reliability and edge cases

  • You've worked with tools like Playwright, Appium, PyAutoGUI, Hammerspoon, or similar

Bonus

Experience with screen reader internals, remote desktop protocols (RDP/VNC), game automation, LLM agent tool-use systems, or mobile device automation (iOS UIAutomation / XCTest, Android UIAutomator / Accessibility).

You've read the whole posting — now see how you match it.