CareerScanCareerScan
JobsCompanies
BlogContact
For Employers
Sign InRegister Free
CareerScanCareerScan

India's verified job platform connecting candidates directly with employers. 100% free applications with instant ATS resume scoring.

Chennai, Bengaluru & Hyderabad
Jobs by location
Jobs in ChennaiJobs in BengaluruJobs in HyderabadJobs in PuneJobs in Mumbai
Popular roles
AR Caller JobsHealthcare Medical CodingReact / Full Stack DeveloperData & Power BI AnalystCustomer Support Executive
Top companies
TCS CareersCognizant JobsInfosys OpeningsApollo HospitalsOmega Healthcare
Career services
Free ATS Resume CheckerAI Resume Builder (Free)AI Job MatcherSalary Guide & BenchmarksJob Alerts on WhatsApp
© 2026 CareerScan India. All rights reserved.256-bit SSL encrypted & verified
Back to all jobs
  1. Home
  2. Jobs
  3. Staff Software Development Test Engineer - AI Evaluation
Tekion
Tekion

Staff Software Development Test Engineer - AI Evaluation

Bangalore HQ
8+ years exp
Full-time
Posted 5d ago
1 views
Actively Hiring Direct 1-Click Apply

Check Your Resume Match Score

Scan your resume against ATS criteria for this Staff Software Development Test Engineer - AI Evaluation role at Tekion.

Apply for this position

Apply on Company Website
Notice a broken link or wrong info?

Job Description

  • About Tekion:

Positively disrupting an industry that has not seen any innovation in over 50 years, Tekion has challenged the paradigm with the first and fastest cloud-native automotive platform that includes the revolutionary Automotive Retail Cloud (ARC) for retailers, Automotive Enterprise Cloud (AEC) for manufacturers and other large automotive enterprises and Automotive Partner Cloud (APC) for technology and industry partners. Tekion connects the entire spectrum of the automotive retail ecosystem through seamless platform. The transformative platform uses cutting-edge technology, big data, machine learning, and AI to seamlessly bring together OEMs, retailers/dealers and consumers. With its highly configurable integration and greater customer engagement capabilities, Tekion is enabling the best automotive retail experiences ever. Tekion employs close to 3,000 people across North America, Asia and Europe.

About the Role

We are looking for a highly motivated Staff SDET – AI Evaluation to join Tekion’s AI Platform team. Evaluation is the backbone of trustworthy AI: as Tekion scales from a handful of AI agents
to 100+ across Service, Sales, F&I, and Analytics, this role builds the evaluation platform and frameworks that let every ML team measure, trust, and improve the quality of AI outputs.In this role, you will be responsible for defining and building Tekion’s AI evaluation capabilities as a shared platform service. You will work closely with ML Engineers, Data Scientists, the AI Platform team, and Product Management to design evaluation datasets, automated scoring pipelines, and quality metrics that quantify the accuracy, consistency, and safety of AIgenerated outputs across the organization. You will own the systems that answer “is this model or agent good enough to ship, and is it
staying good in production?” — from offline benchmarks and LLM-as-judge pipelines to evaluation and continuous quality monitoring. You will also use AI and LLMs to scale evaluation itself, building automated judges and synthetic datasets that expand coverage faster than manual review ever could.

What You’ll Do

  • Develop a deep understanding of Tekion’s AI agents, ML models, and the quality dimensions that matter for each business domain.
  • Design, enhance, and own Tekion’s AI evaluation infrastructure as a shared capability used across ML teams.
  • Create, curate, and maintain evaluation datasets (evals) and golden/ground-truth sets across use cases and domains.
  • Define quality metrics for AI outputs — accuracy, relevance, faithfulness/groundedness, consistency, safety, and task success.
  • Build automated scoring pipelines, including LLM-as-judge, rubric-based, and referencebased evaluation methods.
  • Validate user intents and measure response accuracy and consistency for AI-powered capabilities such as the Analytics Agent.
  • Identify hallucinations, unsafe or biased outputs, and edge cases; design targeted eval suites to catch them.
  • Build both offline evaluation (pre-release benchmarking) and evaluation (production quality monitoring, A/B, drift detection).
  • Establish evaluation gates in CI/CD so model, prompt, or data changes are quality-checkedbefore release.
  • Develop dashboards and reporting that make AI quality visible and actionable for ML and product teams.
  • Use AI/LLMs to scale evaluation — automated judges, synthetic data generation, and eval \tooling
  • Champion evaluation and responsible-AI quality best practices across the organization.

What We’re Looking For

  • 8+ years in SDET, quality engineering, ML engineering, or data science, with hands-on experience building evaluation or measurement systems — or a strong SDET background with deep LLM/ML fluency.
  • Strong programming skills in Python, with the ability to build robust, reusable evaluation pipelines and tooling.
  • Deep understanding of ML/LLM evaluation, benchmark design, and the pitfalls of evaluating non-deterministic systems.
  • Hands-on experience with LLM-as-judge, rubric-based scoring, or human-in-the-loop evaluation.
  • Solid grasp of LLM/agent concepts — prompting, RAG, embeddings, tool use — and generative failure modes (hallucination, drift, prompt sensitivity, bias).
  • Experience designing and curating datasets, including labeling/annotation strategy and data quality.
  • Strong statistical intuition for interpreting eval results and significance.
  • Excellent communication skills to translate quality signals into decisions for ML and product teams.

Nice to Have

  • Experience with eval frameworks/tools such as Ragas, DeepEval, LangSmith, TruLens, Promptfoo, HELM, or provider eval suites.
  • Experience building evaluation, guardrails, or production model monitoring.
  • Familiarity with responsible AI / safety evaluation and red-teaming.
  • Experience with experiment tracking (MLflow, Weights & Biases) and A/B testing.
  • Prior work standing up evaluation as a platform capability for multiple teams

Current Tekion Employees: Please apply via the Internal Job Board in Ashby

Note: Tekion recently transitioned to a new recruiting tool and we appreciate your patience and feedback as we adjust to our new system!

Tekion is proud to be an Equal Employment Opportunity employer. We do not discriminate based upon race, religion, color, national origin, gender (including pregnancy, childbirth, or related medical conditions), sexual orientation, gender identity, gender expression, age, status as a protected veteran, status as an individual with a disability, victim of violence or having a family member who is a victim of violence, the intersectionality of two or more protected categories, or other applicable legally protected characteristics.

For more information on our privacy practices, please refer to our Applicant Privacy Notice here.

Key Requirements & Skills

  • 8+ years in SDET, quality engineering, ML engineering, or data science, with hands-on experience building evaluation or measurement systems — or a strong SDET background with deep LLM/ML fluency.
  • Strong programming skills in Python, with the ability to build robust, reusable evaluation pipelines and tooling.
  • Deep understanding of ML/LLM evaluation, benchmark design, and the pitfalls of evaluating non-deterministic systems.
  • Hands-on experience with LLM-as-judge, rubric-based scoring, or human-in-the-loop evaluation.
  • Solid grasp of LLM/agent concepts — prompting, RAG, embeddings, tool use — and generative failure modes (hallucination, drift, prompt sensitivity, bias).
  • Experience designing and curating datasets, including labeling/annotation strategy and data quality.
  • Strong statistical intuition for interpreting eval results and significance.
  • Excellent communication skills to translate quality signals into decisions for ML and product teams.

Benefits & Perks

medical conditions), sexual orientation, gender identity, gender expression, age, status as a protected veteran, status as an individual with a disability, victim of violence or having a family member who is a victim of violence, the intersectionality of two or more protected categories, or other applicable legally protected characteristics.

Frequently Asked Questions

How to apply for Staff Software Development Test Engineer - AI Evaluation at Tekion?

Click the "Apply on Company Website" button on this page to submit your application directly on the employer's official portal.

What is the salary for this role?

Salary details will be discussed during the interview.

What experience is required?

8+ years of experience is required.

Is this position still open?

Yes, currently active and accepting applications.

ApplicationActively Hiring
Apply on Company Website
Broken link or expired?
Tekion

Tekion

About Tekion | Our Company, Vision & Industry Leadership --> --> --> Tekion One 2026 - Now On-Demand Watch Now Products Company About Us Careers Legal Security & Compliance Blog News Testimonials Learning Security & Compliance Blog Resources Events Tekion AI Agents T1 Accounting AI CRM AI Service AI Login Request Demo Dealers / Retailers DMS CRM Advanced Analytics Digital Retail Digital Service Experience Tekion Pay Tekion Payroll Virtual-to-Visit Experiences Manufacturers / Enterprise Technology Partners Menu Tekion AI Agents T1 Accounting AI CRM AI Service AI Products Automotive

Visit Company Website

More jobs at Tekion

Staff Engineer, DevOps - DevsecOps

Bangalore HQ

Product Counsel

Virtual - California

Customer Value Manager - Ontario

Virtual - Canada

Share this Opening

Job Alerts for qa

Receive email alerts whenever new qa roles in Bangalore HQ are posted.

Set Free Alert →

Similar Openings

Explore related active roles in qa

View all
UrgentActively Hiring
Ashby
QA Engineer
Ashby Verified
10+ years
₹6.2L – ₹8L/mo
Remote - US
qaFull-timeRemote
Posted 1d ago
Apply Now
UrgentActively Hiring
External
Test Engineer - Performance QA
External Verified
5-8 years
Salary not disclosed
India - Bangalore
qaFull-time
Posted 1d ago
Apply Now
Actively Hiring
Finc
Senior QA Engineer
Finc Verified
5+ years
Salary not disclosed
Pune
qaFull-time
Posted 1d ago
Apply Now

Staff Software Development Test Engineer - AI Evaluation

Tekion · Bangalore HQ

Apply on Company Website