Skip to content
aitrainer.work - AI Training Jobs Platform
Software Engineering turing

Senior Python Developer

Turing • Remote, Bangladesh, Brazil, Egypt, India, Kenya, Mexico, Nigeria, Pakistan, Turkey, United States, China

Education

Master's

Type

Hourly

Pay Rate

$30–$65/hr

Our estimate. Turing doesn't publish a rate for this role. How we estimate

Listed

186d ago

Apply opens Turing in a new tab.

Apply Now →

What We Know About This Role

Weekly hours
40 hrs/week
Timezone overlap
Partial overlap with US Pacific time required

About this role

From the Turing listing

About Turing

Turing is one of the world’s fastest-growing AI companies accelerating the advancement and deployment of powerful AI systems. Turing helps customers in two ways: Working with the world’s leading AI labs to advance frontier model capabilities in thinking, reasoning, coding, agentic behavior, multimodality, multilinguality, STEM and frontier knowledge; and leveraging that work to build real-world AI systems that solve mission-critical priorities for companies.Role Overview: This position is within a project with one of the foundational LLM companies. The goal is to assist these foundational LLM companies in enhancing their Large Language Models. One way we help these companies improve their models is by providing them with high-quality proprietary data. This data serves two main purposes: first, as a basis for fine-tuning their models, and second, as an evaluation set to benchmark the performance of their models or competitor models. For example, for SFT data generation, you might have to put together or be provided a prompt which contains provided code and questions, you will then provide the model responses, and write corresponding Python code to solve the questions. For RLHF data generation, you may need to create a prompt yourself or use one provided by the customer, ask the model questions, and evaluate the outputs generated by two versions of the LLM. You'll compare these outputs and provide feedback, which is then used to fine-tune the models. Please note that this role does not involve building or fine-tuning LLMs.

What does day-to-day look like

  • Design, develop, and maintain efficient, high-quality Python code to train and optimize AI models.
  • Conduct evaluations (Evals) to benchmark model performance and analyze results for continuous improvement.
  • Familiarity with Python frameworks and libraries
  • Evaluate and rank AI model responses to user queries across diverse domains, ensuring alignment with predefined criteria.
  • Develop comprehensive explanations and rationales for evaluations, showcasing excellent reasoning and technical expertise.
  • Lead efforts in Supervised Fine-Tuning (SFT), including creating and maintaining high-quality, task-specific datasets.
  • Collaborate with researchers and annotators to execute Reinforcement Learning with Human Feedback (RLHF) and refine reward models.
  • Design innovative evaluation strategies and processes to improve the model's alignment with user needs and ethical guidelines.
  • Create and refine optimal responses to improve AI performance, emphasizing clarity, relevance, and technical accuracy.
  • Conduct thorough peer reviews of code and documentation, providing constructive feedback and identifying areas for improvement.
  • Collaborate with cross-functional teams to improve model performance and contribute to product enhancements.
  • Continuously explore and integrate new tools, techniques, and methodologies to enhance AI training processes.

Requirements

  • 3+ years of strong experience with Python programming language.
  • Industry experience and knowledge of code quality, formatting, and best practices of software development
  • Experience with Python’s testing ecosystem, including unit, integration, and property-based testing.
  • Knowledge of multi-threading and asynchronous programming in Python.
  • Ability to work with architectural patterns and refactor code without introducing regressions.
  • Strong debugging skills, including fixing memory and concurrency issues.
  • Fluent in conversational and written English communication skills

Perks of Freelancing With Turing

  • Work in a fully remote environment.
  • Opportunity to work on cutting-edge AI projects with leading LLM companies.

Offer Details

  • Commitments Required: at least 4 hours per day and minimum 20 hours per week with overlap of 4 hours with PST. (We have 3 options of time commitment: 20 hrs/week, 30 hrs/week or 40 hrs/week)
  • Engagement  type  : Contractor assignment (no medical/paid leave)
  • Duration of contract : 1 month; [expected start date is next week]

Evaluation Process (approximately 75 mins)

  • Two rounds of interviews (60 min technical + 15 min cultural &, offer discussion)

How long hiring takes

Across the AI training platforms we refer candidates to, the median gap between referral and hire is about 30 days. It varies by platform and role, so treat it as a rough guide for this one.

Within 2 weeks
~25%
Within 6 weeks
~60%
Within 3 months
~80%

Interview Prep

Sample questions for a Python Developer role, written in-house to help you prepare.

How do you decide between using a list, a generator, or a NumPy array when processing a large dataset?

I use a generator when I only need to iterate once and don't need random access, since it keeps memory flat regardless of dataset size. I reach for NumPy arrays when doing numeric operations at scale, since vectorized operations are far faster than looping over a plain list.

What causes a memory leak in a long-running Python process, and how do you find one?

Common causes include holding references in a growing global cache, circular references involving objects with custom __del__ methods, or accumulating data in a list that's never cleared. I use tools like tracemalloc or objgraph to snapshot memory over time and identify which object types are growing unexpectedly.

How do you approach debugging a function that works correctly in isolation but fails when called from a larger pipeline?

I check whether shared mutable state, like a list or dictionary passed by reference, is being modified somewhere upstream before the function receives it. I also verify the actual arguments being passed at the call site with a debugger rather than assuming they match my mental model.

What's your approach to writing tests for code that depends on an external API?

I mock the external call at the boundary so tests run deterministically and don't depend on network availability, and I write a smaller set of integration tests that hit the real API to catch contract drift. Relying only on mocks risks tests passing while the real integration is broken.

See all 10 questions for this role →

This listing calls for this tool directly. Prep for the technical screen:

What to Expect

Looking at Turing Software Engineering listings we've tracked, contracts in this domain typically run about 8.4 weeks. Actual length varies by project, but this gives you a realistic baseline going in.

Based on 28 extracted Turing Software Engineering listings.

Why this role

This Senior Python Developer role works inside a project for a foundational LLM company, applying Python development skills to help train and evaluate frontier models. It's open across a wide pool including Bangladesh, Brazil, Egypt, India, Kenya, Mexico, Nigeria, Pakistan, Turkey, the USA, and China.

Skills and categories

Explore other opportunities in related specializations:

Related jobs

Turing

Browse All Jobs from Turing

Discover more opportunities on Turing that match your skills and interests.

View All Turing Jobs →

Verified Reviews

Loading reviews…

Community Reviews

Loading reviews…
💬

Share your experience with Turing

Help other candidates make better decisions by leaving a review.

Sign in to leave a review

Common questions

Do I need to be a software engineer to work for Turing?

No, not anymore. Turing built its name matching senior engineers with Silicon Valley companies, but it has since expanded into AGI infrastructure work and now hires non-engineering domain experts, technical writers, and researchers for post-training data annotation and RLHF. A strong analytical background and excellent English matter more than coding ability.

How does Turing's talent matching work?

Turing calls it the Intelligent Talent Cloud. You build a profile and go through vetting (automated tests, an AI-powered interview, practical skill assessments), and once vetted, Turing's algorithm surfaces your profile directly to partner companies like Fortune 500s and top AI labs. You don't browse listings or bid on work; matches come to you.

What does asynchronous AI training work mean in practice?

No set hours, no check-ins, no meetings. You log in when you want, pick up an available task, complete it, and submit; nobody is waiting on you in real time. That's different from remote employment, where you're expected online during business hours. The tradeoff: you're competing with others for available tasks, so an empty queue means there's simply nothing to do until more work is released.

What does Software Engineering work look like for a Senior Python Developer?

Tasks here are scoped to Software Engineering, not generic labeling. As a Senior Python Developer, expect to draw on real domain judgment (evaluating outputs, correcting errors, or providing expert reasoning specific to Software Engineering) rather than following a one-size-fits-all rubric. If you don't have hands-on Software Engineering background, this is likely not the right listing to start with.

What specific skills does this listing call for?

Critical Thinking, Business Analysis, Coding, and Python are named directly in the listing. If you don't have hands-on experience with these, expect the screening process to test for them directly rather than accepting adjacent experience as a substitute.

How many hours per week does this role require?

Based on the listing, this role is scoped at about 40 hours per week. Treat this as a real commitment expectation, not a loose estimate.

How much does this specific role pay?

The listing doesn't state a rate. The $30–$65/hr shown here is our estimate from the role type and location (see /pay-methodology), so treat it as a rough guide and confirm the actual rate with the platform before committing time.

What happens when I click Apply on this listing?

You'll be taken to Turing's external site to complete your application there. This listing links through a referral, but the process is identical to applying directly; the link just routes you correctly. Create an account on their site and follow their onboarding steps.

Do I need a Master's to qualify?

This role lists a master's degree as a requirement. In practice, the domain assessment is the real gate. If you can pass it, the degree is usually secondary. However, some platforms verify credentials formally, so list your actual qualifications accurately on your profile.

Can I apply from outside Bangladesh, Brazil, Egypt and 8 other countries?

This specific role is open only to people based in Bangladesh, Brazil, Egypt, India, Kenya, Mexico, Nigeria, Pakistan, Turkey, the United States, and China. If you are somewhere else, applying is unlikely to lead to an offer even if you pass the assessment, because the restriction is usually about where the work can legally be contracted rather than your skills. Read the full description for any tax-residency or right-to-work caveats before you apply, since they can differ by country.