Filters
Create alert
Sort by
  • Relevance
  • Date
Exact location
  • Auto
  • Exact location
  • Less than 15 km
  • Less than 25 km
  • Less than 35 km
  • Less than 45 km
  • Less than 55 km
  • Less than 65 km
  • Less than 75 km
Company
  • Accenture
  • Amazon
  • AWS
  • Intuit
  • Trianz

Judge Jobs in Bangalore

1 - 15 of 29
1 - 15 of 29
Search Results - Judge Jobs in Bangalore
apartmentSystems LimitedplaceBangaloreevent_available
in terms they can act on Train delivery teams on AI-specific testing practices Own the evals framework for the practice — golden datasets, scoring rubrics, LLM-as-judge calibration, and versioned benchmarks per use case Define eval acceptance thresholds per...
apartmentEmeritusplaceBangaloreevent_available
world-class education from top tier universities Wharton, Columbia, INSEAD, MIT Sloan, Berkeley, Kellogg, Cambridge-Judge & Harvard to senior and top management executives can help launch them into a high growth career.  •  Build and manage pipeline...
apartmentK2 Partnering SolutionsplaceBangaloreevent_available
Cortex Search and Cortex Analyst, or demonstrable equivalent Retrieval evaluation: golden sets, LLM-as-judge, hallucination rate, attribution Terraform and SDLC literacy Owns Coordinator hardening, Critic agent, guardrails Chunking strategy and the chunk...
apartmentTricog HealthplaceBangaloreevent_available
multiple open requests at once  •  A basic eye for design/creative quality (you don't need to design, but you should be able to judge if something looks right)  •  Comfortable with MS Office/Google Workspace, Excel; familiarity with tools like Asana, Notion...
apartmentDentsu Global ServicesplaceBangaloreevent_available
team. Desirable / Nice-to-Have Skills Experience with some of the following is beneficial but not required: Hands-on experience evaluating GenAI/LLM applications (golden sets, LLM-as-judge, eval frameworks such as Promptfoo, DeepEval, Ragas, or similar...
apartmentNeoplaceBangaloreevent_available
workflows.  •  Evaluation & experimentation — Golden sets, offline metrics (NDCG, MRR, recall@k), calibrated LLM-as-judge pipelines, and online A/B testing; regression gating as a habit, not an afterthought.  •  Search infrastructure at scale — Internals...
apartmentDentsu Global ServicesplaceBangaloreevent_available
experience evaluating GenAI/LLM applications (golden sets, LLM-as-judge, eval frameworks such as Promptfoo, DeepEval, Ragas, or similar). /n  •  Experience with Google Vertex AI, Gemini, or a comparable LLM platform. /n  •  Exposure to MCPs, or tool-use...
apartmentKakeplaceBangaloreevent_available
the engineer who judges correctness.  •  Experience with e-commerce, payments, video platforms, or other correctness-critical domains is a strong plus.  •  Availability to work aligned with the partner's time zone. Additional  •  US Timezone Overlap: 5h–6h daily...
apartmentTricog HealthplaceBangaloreevent_available
/n  •  A basic eye for design/creative quality (you don't need to design, but you should be able to judge if something /n /n looks right) /n  •  Comfortable with MS Office/Google Workspace, Excel; familiarity with tools like Asana, Notion, or similar...
apartmentTekion CorpplaceBangaloreevent_available
frameworks — LangChain/LangGraph, LlamaIndex, CrewAI, OpenAI Agents SDK, or similar orchestration tools  •  Experience with LLM observability and evaluation — tracing (LangSmith, OpenTelemetry), LLM-as-judge evaluation, cost and latency monitoring...
apartmentTrianzplaceBangaloreevent_available
such as instruction tuning, LoRA, QLoRA, DPO, RLHF, and related methodologies. /n  •  Build and optimize LLM-as-a-judge evaluation systems and benchmarking frameworks. /n  •  Drive model performance improvements through experimentation, testing, and continuous refinement...
apartmentTekion CorpplaceBangaloreevent_available
orchestration tools  •  Experience with LLM observability and evaluation — tracing (LangSmith, OpenTelemetry), LLM-as-judge evaluation, cost and latency monitoring  •  Understanding of AI safety and guardrails — input/output validation, PII detection, content...
apartmentAccentureplaceBangalorelanguageaccenture.comevent_available
LangChain, LlamaIndex, AutoGen, or equivalent Fine-tuning experience — LoRA, QLoRA, PEFT dataset curation and evaluation. – vry good to have LLM evaluation frameworks — RAGAS, LLM-as-judge, regression testing, quality gates. Cloud & infrastructure Deep...
apartmentnAbleplaceBangalorelanguagefoundit.inevent_available
synonyms) as the shared vocabulary AI agents use to answer business questions.  •  Stand up the evaluation harness for the semantic and agentic layer, including golden-question datasets, LLM-as-a-judge scoring, and side-by-side runs that gate every release...
apartmentAmazonplaceBangalorelanguageamazon.jobsevent_available
tools to support your work across non-sensitive workstreams. You will work alongside AI automation (LLM judges, AI agents) as a human-in-the-loop to ensure quality and safety standards are met. Associates are expected to use AI-powered tools (e.g., LLM...
12

Companies now hiring in Bangalore:

Judge jobs – More locations:

Broaden your job search:

Don’t miss out on new job vacancies!
Create a job alert for: Judge, Bangalore
It's free, and you can cancel email updates at any time
12
Get new jobs by email!
Get email updates for the latest Judge jobs in Bangalore
It's free, and you can cancel email updates at any time