AI Benchmark Engineer | Native Language Specialist - French (France)
LILT · AI/ML
Source: Arbeitnow
No hybrid days in this listing.
Checked for remote eligibility. This role is open to applicants based in Hungary.
How this role compares to the AI/ML market
Across all remote AI/ML roles on RemoteJobSearch. Refreshed every 30 minutes.
63 of 263 AI/ML roles publish a salary
Median of published salaries. Full reported spread in this market: $75,000–$351,000, outliers included.
22% of these roles
LILT is seeking a French native software engineer to design and validate LLM multilingual benchmarks in a 100% remote, freelance capacity.
This role involves creating complex task environments for the Terminal-Bench evaluation suite, focusing on non-English data processing and encoding edge cases. The position requires senior-level technical expertise in Python and CLI workflows to stress-test large language models. While the schedule is flexible, candidates are expected to commit at least 10 hours per week to maintain project consistency.
Why You’ll Love This Role
What You’ll Bring
What You’ll Get
About the Company
LILT provides multilingual AI and human-verified services to enterprises, governments, and AI developers. They specialize in language technology and global communication infrastructure.