Skip to main content
Loading
loadingbar
Loading, Please wait..!!

Remote Senior Python Engineer – LLM Evaluation (US-based)

  • Job type Posted on: May 25, 2026
  • Experience level Turing
  • Employment type Denver, Colorado
  • Employment type Remote
  • Salary Full-time

Point Apply Here APPLY LATER

Curious about compensation?

Explore the historical salary trends, average pay, and estimated compensation for Remote Senior Python Engineer – LLM Evaluation (US-based) roles in Colorado.

View Salary Guide →

Job Title :

Remote Senior Python Engineer – LLM Evaluation (US-based)

Job Type :

Full-time

Job Location :

Denver Colorado United States

Remote :

Yes

Jobcon Logo Job Description :

Based in San Francisco, California, Turing is the world’s leading research accelerator for frontier AI labs and a trusted partner for global enterprises deploying advanced AI systems. Turing supports customers in two ways: first, by accelerating frontier research with high-quality data, advanced training pipelines, plus top AI researchers who specialize in software engineering, logical reasoning, STEM, multilinguality, multimodality, and agents; and second, by applying that expertise to help enterprises transform AI from proof of concept into proprietary intelligence with systems that perform reliably, deliver measurable impact, and drive lasting results on the P&L. Ideal Background This role is ideal for engineers who have worked at the frontier of AI — at companies like OpenAI, NVIDIA, Databricks, Palantir, Snowflake, or similar organizations pushing the boundaries of intelligent systems. We especially welcome graduates from top computer science programs such as Stanford, MIT, Carnegie Mellon, UC Berkeley, Georgia Tech, and comparable institutions — though exceptional experience and skill always take precedence over pedigree. Responsibilities Evaluate and refine AI-generated code across backend and frontend contexts to ensure that it is efficient, scalable, and reliable. Collaborate with cross‑functional teams to enhance AI-driven coding solutions against industry performance benchmarks. Build agents that can verify the quality of the code and identify error patterns across full‑stack applications. Hypothesize on steps in the software engineering cycle (prototyping, architecture design, API design, production implementation, launch, experiments, monitoring, operational maintenance) and evaluate model capabilities on them. Design verification mechanisms that can automatically verify a solution to a software engineering task. Required Skills Several years of software engineering experience (3 years or more) Experience deploying scalable, production‑grade software using modern languages and tools. Deep understanding of software architecture, design, development, debugging, and code quality/review assessment. Excellent oral and written communication skills for clear, structured evaluation rationales. Commitment: flexible engagement, minimum 10 hrs/week, up to 40 hrs/week Type: Contractor (no medical/paid leave) Duration: 1 month (potential extensions based on performance and fit) Location: Candidates must be based in the United States #J-18808-Ljbffr

Jobcon Logo Position Details

Posted:

May 25, 2026

Reference Number:

14660_78A1E063E89B4B1DE40974F43BF4D2FA

Employment:

Full-time

Salary:

Not Available

City:

Denver

Job Origin:

APPCAST_CPC

Share this job:

  • linkedin

Jobcon Logo
A job sourcing event
In Dallas Fort Worth
Aug 19, 2017 9am-6pm
All job seekers welcome!

Remote Senior Python Engineer – LLM Evaluation (US-based)    Apply

Click on the below icons to share this job to Linkedin, Twitter!

Based in San Francisco, California, Turing is the world’s leading research accelerator for frontier AI labs and a trusted partner for global enterprises deploying advanced AI systems. Turing supports customers in two ways: first, by accelerating frontier research with high-quality data, advanced training pipelines, plus top AI researchers who specialize in software engineering, logical reasoning, STEM, multilinguality, multimodality, and agents; and second, by applying that expertise to help enterprises transform AI from proof of concept into proprietary intelligence with systems that perform reliably, deliver measurable impact, and drive lasting results on the P&L. Ideal Background This role is ideal for engineers who have worked at the frontier of AI — at companies like OpenAI, NVIDIA, Databricks, Palantir, Snowflake, or similar organizations pushing the boundaries of intelligent systems. We especially welcome graduates from top computer science programs such as Stanford, MIT, Carnegie Mellon, UC Berkeley, Georgia Tech, and comparable institutions — though exceptional experience and skill always take precedence over pedigree. Responsibilities Evaluate and refine AI-generated code across backend and frontend contexts to ensure that it is efficient, scalable, and reliable. Collaborate with cross‑functional teams to enhance AI-driven coding solutions against industry performance benchmarks. Build agents that can verify the quality of the code and identify error patterns across full‑stack applications. Hypothesize on steps in the software engineering cycle (prototyping, architecture design, API design, production implementation, launch, experiments, monitoring, operational maintenance) and evaluate model capabilities on them. Design verification mechanisms that can automatically verify a solution to a software engineering task. Required Skills Several years of software engineering experience (3 years or more) Experience deploying scalable, production‑grade software using modern languages and tools. Deep understanding of software architecture, design, development, debugging, and code quality/review assessment. Excellent oral and written communication skills for clear, structured evaluation rationales. Commitment: flexible engagement, minimum 10 hrs/week, up to 40 hrs/week Type: Contractor (no medical/paid leave) Duration: 1 month (potential extensions based on performance and fit) Location: Candidates must be based in the United States #J-18808-Ljbffr

Loading
Please wait..!!