Senior MLOps Software Engineer
Tata Consultancy Services Warsaw, Mazowieckie, Poland Senior
Wynagrodzenie do uzgodnienia
🪄 Prompt EngineeringStacjonarnieB2B CONTRACT
Aplikuj na tę ofertę
Wyślemy Twój profil bezpośrednio do firmy.
O roli
Your daily tasks:
- Design, build, and own the software solution used to assess cyber capabilities in AI models.
- Define a measurable benchmark and evaluation methodology covering relevant cyber-capability scenarios, tasks, scoring criteria, thresholds, and controls.
- Implement repeatable, automated evaluation pipelines that execute benchmark tests and generate consistent model-level results.
- Analyze model outputs, quantify performance, investigate failure modes, and translate findings into clear technical evidence.
- Partner with Model Risk Management, cyber security, data science, engineering, and governance stakeholders to align the approach with review expectations.
- Document evaluation assumptions, scope, data sources, test cases, metrics, limitations, versioning, and decision rationale to support independent challenge and auditability.
- Build reliable software components, APIs, data-processing workflows, and reporting outputs using secure engineering practices.
- Integrate evaluations into CI/CD and MLOps workflows, including experiment tracking, reproducibility, monitoring, and controlled release processes.
- Maintain benchmark relevance as model capabilities, cyber risks, and evaluation techniques evolve.
- Provide technical leadership, conduct code and design reviews, and drive the solution from proof of concept through production ownership.
Your attitudes we value
- 7+ years of professional experience in software engineering, MLOps, machine learning engineering, data engineering, or a closely related discipline.
- Strong software engineering skills, preferably with Python, including modular design, testing, version control, APIs, and production-quality coding standards.
- Hands-on experience evaluating machine learning or generative AI model performance using quantitative and qualitative methods.
- Strong data analysis skills, including metric design, experiment analysis, error analysis, statistical reasoning, and clear visualization or reporting of results.
- Practical experience developing automated model evaluation or benchmarking frameworks and converting evaluation outcomes into decision-ready evidence.
- Experience interfacing with Model Risk Management, independent validation, risk, compliance, or governance teams.
- Working knowledge of MLOps practices such as CI/CD, model and dataset versioning, experiment tracking, reproducible execution, monitoring, and deployment controls.
- Ability to own a technical solution end to end, from requirements and architecture through implementation, validation, documentation, deployment, and support.
- Strong written and verbal communication skills, with the ability to explain complex evaluation results to technical, risk, and non-technical stakeholders.
What we offer:
- Hybrid working model (2 days in the office from our Warsaw office, 3 days remote per week)
- Competitive salary with annual salary reviews
- Opportunity to use foreign languages and the newest technologies on a daily basis
- Indefinite period employment contract after probation period
- Access to MyBenefit platform where you can choose from a variety of benefits (Sport, Tourism, Culture, Recreation etc.)
- Sport's card (Multisport)
- Private medical care for you and your family
- Private life insurance
- Access to wide range of learning & development platforms
- Discounted theatre tickets
- Work in a multicultural environment