Agile Robots SE
Senior C++ Performance Engineer (all genders)
Munich, Bavaria, Germany
Get past the screening software and onto a recruiter's desk
hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.
- Keywords matched to this posting
- Fit score before you apply
- Cover letter included
Apply from your AI assistant
Connect hirly to Claude and ask it to apply to this job. hirly tailors your resume, fills the employer’s form and asks before sending. ChatGPT: manual setup today.
Some employer sites stop an application at a CAPTCHA or sign-in and hand it back with a link. Applying needs a paid plan. Works with any assistant that supports MCP.
hirly's read of this role
- Seniority
- Senior
- Country
- DE
- Work mode
- On-site / unstated
- First seen by hirly
- 8 Oct 2026
Derived automatically from the posting. Upload your resume above to see how the role scores against it.
the posting
About the role
At Agile Robots, we build powerful software frameworks to orchestrate complex robotic systems and solve challenging tasks. The solutions are sold to various international customers to enable them to control their robotic work cells. Smooth integration and scalable deployment are key goals.
AgileCore is meant to be a scalable platform that supports multiple projects in parallel. That only works if the platform underneath scales with it - if it stays fast, thread-safe, and diagnosable as projects, devices, and operating conditions multiply.
We are looking for a Senior C++ Performance Engineer to join our AgileCore Runtime core team. This is a T-shaped role in a small senior team: you are a full member of the core team first - contributing to the Runtime, reviewing contributions from the teams that build on our platform, and sharing our support and release duties - and within the team, you own the performance and stability practice as your specialization.
On the specialization side, you will drive the performance of the Runtime core - the Blackboard, the Behavior Engine and the execution paths beneath them - make the system provably thread-safe under load, and make failures in the field fast to diagnose. You own the practice: the methodology, the tooling, the benchmarks, the audits and the education. You work hand in hand with the platform architect, who owns the performance budgets and hot-path interface contracts at the architecture level - you turn those budgets into measured, enforced reality and shape them with your findings. You set the standard for performance, concurrency and troubleshooting, and raise the team's ability to meet it without slowing feature development down. This is a shipping role, not a research role: success is measured in merged improvements, closed field escalations and green regression gates - not in microbenchmark leaderboards. We optimize to the budget the product needs, not to zero.
Why this role
Greenfield performance infrastructure: the benchmark suites, budgets and CI regression gates are yours to build, not inherit
A measurable, physical problem: your profiling and your benchmarks run against real robots on our hardware-in-the-loop rigs - where third-party plugins meet 1 kHz vendor control loops
A small senior core team with real ownership: your work becomes the standard the whole platform builds on
Your Responsibilities:
Own the latency, throughput, and jitter of the Runtime core components - the Blackboard (our typed, subscribable in-process data plane) and the Behavior Engine (skill execution tree, trigger handling) - and carry improvements end to end: from flame graph to merged change to a regression gate that keeps it fixed
Build the real-time basis for the platform: scheduling policies and priorities, CPU affinity, allocation, and page-fault discipline on hot paths, and lock-free structures where they pay for themselves
Be a full member of the Runtime core team: develop core components, review contributions from feature teams building on the platform, and take part in the team's support and release rotations
Make performance measurable and enforced - benchmark suites, per-subsystem budgets, and regression gates in CI instead of surprises at a customer site
Be the hands-on authority on thread safety: ownership models, lock hierarchies, memory-order reasoning, and eliminating whole classes of races through sanitizers, stress and soak testing
Ensure the product holds across the complete operational domain it is used in - every supported device, load profile, and deployment topology, not only the configurations that appear in a demo
Improve debuggability so that factory deployment and support scale: tracing, structured logging, runtime metrics, crash and core-dump handling, symbols for stripped release binaries
Carry performance and stability into the Plugin SDK and its C ABI, where third-party plugins meet vendor control loops running at 1 kHz
Investigate complex issues across software, OS, network, and hardware boundaries - including the hardest escalations from the field
Partner with the platform architect on performance-critical design: propose and implement improvements within the platform's performance budgets and hot-path interface contracts, and feed measurements back into how those budgets are set
Mentor engineers in performance, concurrency, and troubleshooting through code reviews, deep-dive sessions, design discussions and technical guidance - without producing overhead
Essential Skills:
7+ years of professional software engineering experience, with strong hands-on C++ development
Based in Munich or open to relocate - we support relocation and visa sponsorship
Strong written and verbal communication skills in English
Strong experience with modern C++, ideally C++17 or C++20, including templates and compile-time design
Proven performance engineering: profiling before optimizing, reading flame graphs and hardware counters, and reasoning about tail latency and jitter rather than averages
A track record of shipped performance work: improvements you carried from measurement through merged code to sustained enforcement in CI - not benchmarks in isolation
Strong experience with multi-threaded and concurrent programming: memory models, atomics, lock-free patterns, and the judgment to know when a plain mutex is the right answer
Exposure to real-time or near-real-time systems on Linux is strongly preferred: scheduling policies, preemption, priority inversion, CPU affinity, and the usual sources of jitter - deep prior RT experience is a plus, not a gate; the performance fundamentals matter more
Strong debugging skills using GDB, sanitizers (ASan/TSan/UBSan), Valgrind, perf, strace, ltrace, tcpdump, and core dumps
Solid understanding of Linux internals: processes, threads, IPC, signals, scheduling, memory management, buffering, and dynamic linking
Experience with CMake and compiler toolchains (GCC, Clang)
Experience writing automated tests, including unit tests, integration tests, and hardware-related test strategies, as well as benchmarks and stress/soak tests
A pragmatic, team-first senior mindset: you optimize the things that matter to the product and stop when the budget is met; you share reviews, support, and release duties in a small senior team; you make your specialty a capability of the whole team rather than a silo - and you would rather ship a 30% win this week than discuss a 3% win for a month
Beneficial Skills:
Database internals: storage and index structures, concurrency control, MVCC, transaction and cache design - our Blackboard is essentially an in-process typed store
Contributions to high-performance or systems software: allocators, runtimes, engines, databases, network stacks, kernels
A low-latency background: trading systems, real-time audio or video, telecom, or embedded motion control
Extensive template metaprogramming and compile-time computation
Knowledge of real-time Linux, kernel scheduling, IRQ handling, kernel/user-space boundaries, and latency measurement with cyclictest, ftrace or eBPF
C++ ABI and plugin / shared-library API design that stays stable across compilers and toolchain versions
Hands-on experience with instrumentation profilers such as Tracy (we use it), as well as perf, VTune, or similar
Windows internals and the MSVC toolchain - parts of the platform also ship on Windows
Networking fundamentals: TCP/IP, UDP, sockets, latency, buffering, and packet flow
Experience designing maintainable architectures - clear ownership, separation of concerns, testability - in systems constrained by time, not only by features
What we offer
A dynamic high-tech company, combined with financial soundness and world-class investors.
Join an interdisciplinary, international team with 60+ different nationalities in a collaborative work environment.
Lots of devel
Listed on hirly, a job board. hirly is not the employer: Agile Robots SE is hiring for this role.
Similar jobs
- Senior Performance EngineerNvidia · 3 LocationsFirst seen yesterday
- System Performance Engineer (m/f/d)OHB-System AG · Bremen, GermanyFirst seen 4d ago
- Process Performance Engineer (m/w/d) für AIRBUS AerostructuresSimpleXX GmbH · Augsburg, Bayern, Bayern, GermanyFirst seen 7d ago
- Principal Database Performance Engineer - Core Engineering (C++)ClickHouse · GermanyFirst seen 12d agoremote
- Process Performance Engineer (m/w/d)fabplus GmbH · Augsburg, Bayern, Bayern, GermanyFirst seen 13d ago
Browse similar roles
Want this one?
Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.
Tailor my resume for this job