This role offers the opportunity to contribute directly to the evolution of advanced AI systems by improving their reliability, accuracy, and safety.
You will work on evaluating large language models through structured testing, quality analysis, and adversarial scenarios. The position combines AI quality assurance, data evaluation, and technical problem-solving to strengthen model performance. You will design evaluation frameworks, identify failure patterns, and help improve AI reasoning capabilities across diverse use cases. Working with cutting-edge AI technologies, you will influence how future systems handle information, reasoning, and real-world tasks.