Agentic Compiler Engineer
Kog
Curated Remote Listing: This opportunity was indexed from arbeitnow.fr. You can submit your proposal directly to the employer or manage the engagement through your free Clivora workspace with $0 platform deal fees.
Role Overview
to go deep in your strongest technical area while expanding into the other parts of the stack. High ownership over technical decisions and systems that will shape how Kog optimizes LLM inference. This role is based in Paris, and we are looking for candidates who can relocate to Paris and work closely with the team.
Key Responsibilities
- You will work directly on AGCO.
- The goal is to build a system that can explore ways to optimize LLM execution, generate changes, compile them, check correctness, run them on real hardware, measure the results, and use this feedback to guide the next optimization.
- You will contribute to areas such as: Compiler and IR design for representing and transforming LLM computations.
- Optimization passes, lowering, and code generation.
- Search methods for exploring different implementations and execution strategies.
- Verification and correctness checks for generated changes.
- GPU execution, profiling, and performance optimization.
- LLM inference across operators, memory, parallelism, and communication.
- Optimization loops that connect generated changes to measurements on real GPUs.
- One direction we are exploring combines an IR, a verifier, a compiler, and a search optimizer.
- We plan to start with focused problems, build working prototypes, and extend the system from what we learn.
- Your main area will depend on your experience, skills, and interests.
- You may focus more on compilers, GPU systems, or LLM inference while working closely with people across the full stack.
- WHAT WE LOOK FOR We look for engineers with deep technical expertise and original work in at least one area relevant to AGCO.
- Relevant experience includes: Compiler engineering, including optimization passes, IRs, lowering, code generation, LLVM, or MLIR.
- GPU programming with CUDA, HIP, Metal, Vulkan, or similar technologies.
- GPU performance work involving kernels, memory, synchronization, profiling, or hardware behavior.
- LLM inference engines and performance optimization.
- Attention, MoE, parallelism, communication, or other systems-level parts of LLM execution.
- Formal verification, equivalence checking, SAT/SMT, or related methods.
- Systems that generate, search, test, benchmark, or optimize code automatically.
- We care about what you personally built and the technical decisions behind it.
- Strong candidates can explain the problem, their approach, the alternatives they explored, and how they measured the result.
- We review technical work during the process.
- This can be public code, an upstream contribution, a paper, a thesis, a technical project, or a detailed write-up based on work you can share
Benefits & Perks
You will join a small team building AGCO as a core part of Kog's technology. Work at the intersection of compilers, GPU systems, and LLM inference. Direct access to engineers working across the full inference stack. A fast loop from an optimization idea to compilation, execution, verification, and measurement on real GPUs
Apply & Contact
Apply directly with Kog on their official portal, or submit your proposal directly on Clivora with $0 platform take rates.
Founder Platform Guarantee
Founded by Usman Ghias at CODCrafters to free independent developers and agencies from predatory platform commissions.
Hiring for Kog?
Claim this listing to review inbound proposals directly, message candidates with zero intermediary fees, and manage contracts inside Clivora.
Claim this employer profile →