Crea un profilo in modo da poter essere trovato dalle aziende, ottenere offerte di lavoro più adatte alle tue esigenze e candidarti più velocemente.
  • Cerca lavoro
  • Preferiti
  • Crea CV
    Novità
  • Stipendi
  • Iscrizioni

AI Safety Researcher

CTI Clinical Trial and Consulting Services

The work

Icaro Foundation is an independent non-profit AI safety lab based in Rome. We study advanced AI systems: what they can do, how they fail, and how those findings can support developers and institutions responsible for their governance.

We see AI safety as one of the defining scientific and societal challenges of our time. As AI systems become more capable, autonomous, and widely deployed, understanding and reducing their risks is increasingly urgent. We are looking for people who are deeply interested in these questions and motivated to contribute through rigorous research.

You will help produce new research and develop the lab’s shared codebase and knowledge base , working closely with our researchers across the research process: reviewing literature, refining questions, implementing experiments, analysing results, and contributing to papers and technical reports.

Our research focuses particularly on agentic, multi-agent, and compositional safety : how risks emerge across extended interactions, tool use, and systems involving multiple AI agents. We also study testing awareness and evaluation validity , including whether models behave differently when they recognise that they are being evaluated.

Alongside our research, we evaluate frontier models for international model providers as independent third-party evaluators, using public and proprietary benchmarks and red-teaming environments.

Our public work includes:

  • Boiling the Frog, on multi-turn agentic safety;
  • Adversarial Humanities Benchmark, on the robustness of safety behaviour under stylistic reformulations;
  • research on LLM-to-LLM risks, multi-agent collusion, and interaction-level safety.

You can explore our research programme and papers to learn more.

What you would do

Your work will combine three closely connected areas.

Contribute to research

  • Review relevant literature, compare methods, and identify questions worth investigating.
  • Help turn research questions into experimental protocols, including baselines, controls, and clear evaluation criteria.
  • Implement and run experiments with frontier and open-weight models, including agentic and multi-agent environments.
  • Analyse results and model traces, investigate unexpected behaviour, and assess confounders and alternative explanations.
  • Contribute to research papers, benchmarks, technical reports, and presentations.

Develop the research codebase

  • Write and improve Python code for experiments, evaluations, data processing, and analysis.
  • Extend existing tools and environments, fix bugs, and participate in code review.
  • Add tests, documentation, and reproducible configurations so other researchers can inspect, rerun, and build on your work.

Build the lab’s knowledge base

  • Produce concise, source-grounded notes on papers, methods, benchmarks, and research questions.
  • Document experimental setups, findings, limitations, and negative results.
  • Organise and connect references, datasets, code, and research notes so the team can find relevant evidence and reuse previous work.

You may bring stronger skills in research or engineering. The role involves both writing code and reasoning carefully about evidence.

Who should apply

We welcome applications from master’s students, PhD students, recent graduates, and researchers at the beginning of their careers , including those who have recently completed a PhD.

Relevant experience may come from a thesis, academic research, independent experiments, open-source contributions, internships, or previous employment. We also welcome applicants from non-traditional backgrounds who can demonstrate strong research or engineering ability.

A completed PhD, previous AI safety employment, and published papers are not required. We care about the quality of your work, your contribution to it, and your ability to learn.

If you are currently studying, please tell us about your availability and how you would combine the role with your academic commitments.

What we are looking for

  • Experience with LLM evaluations, red-teaming, or benchmark development;
  • A solid technical or quantitative background, developed through university study, independent projects, or relevant work.
  • Good Python skills and familiarity with Git, debugging, and working with an existing codebase.
  • Practical experience with machine learning or LLMs through at least one substantive project or research contribution.
  • An understanding of basic experimental reasoning and statistics: comparing conditions, interpreting results, and recognising uncertainty and possible confounders.
  • The ability to read technical papers critically and explain methods, findings, and limitations clearly in English.
  • A strong interest in AI safety and a sense of urgency about understanding and reducing the risks posed by increasingly capable AI systems.
  • Intellectual curiosity, openness to criticism, and a willingness to revise your views in response to evidence.

Useful, not required

Experience with:

  • agentic or multi-agent systems;
  • statistical analysis or experimental replication;
  • software testing, containers, or reproducible research workflows;
  • literature reviews, research documentation, or open-source contributions;
  • Inspect AI, the open-source evaluation framework developed by the UK AI Security Institute and Meridian Labs, or comparable tools.

For an example of our research software, see the Adversarial Humanities Benchmark codebase, also listed in Inspect Evals as an externally maintained evaluation.

You do not need experience in all of these areas.

How we work

We are a small research team. You will work closely with experienced researchers and receive feedback on experimental design, code, analysis, and writing.

You will begin with clearly scoped contributions to ongoing projects and take on greater responsibility as your skills and familiarity with the work develop. We encourage everyone to ask questions, challenge assumptions, and propose ideas.

Existing evaluation infrastructure, technical support, and API budget are available. Contributions may become public papers, benchmarks, datasets, or tools where compatible with confidentiality obligations. Authorship and acknowledgement will reflect contributions.

We value work that others can understand and build on: clear reasoning, reliable code, well-documented experiments, and honest reporting of uncertainty.

  • Location: Flexible, with a preference for working in person with the team in Rome, Italy .
  • In-person collaboration: We particularly welcome applicants who are based in Rome or would be interested in relocating. We value regular in-person discussion, collaborative experimentation, and learning from one another.
  • Remote arrangements: May be considered for candidates based in Europe or China, with substantial overlap with European working hours.
  • Engagement: Contractor role.
  • Compensation: The specific compensation range will be shared during the first interview.

Referrals increase your chances of interviewing at Icaro Foundation by 2x

Find curated posts and insights for relevant topics all in one place.

Seniority level

  • Mid-Senior level

Employment type

  • Contract

Job function

  • Engineering and Information Technology
  • IT System Testing and Evaluation
#J-18808-Ljbffr
Offerta di lavoro pubblicata 1 giorno fa
Offerte di lavoro simili
  • pbThe work /b /p pIcaro Foundation is an independent non-profit AI safety lab based in Rome. We study advanced AI systems: what they can...  ...these questions and motivated to contribute through rigorous research. /p pYou will bhelp produce new research and develop the lab’s... 
    Consigliato
    Stage/Tirocinio
    Orario flessibile

    CTI Clinical Trial and Consulting Services

    Roma
    1 giorno fa
  •  ...Icaro Foundation in Rome seeks a mid- to senior-level researcher to advance safety research for frontier AI systems. You will review literature, design experiments, implement code, and contribute to papers and technical reports. The role blends research and engineering... 
    Consigliato
    Remoto

    CTI Clinical Trial and Consulting Services

    Roma
    1 giorno fa
  •  ...PwC Italy cerca un AI Developer per unirsi al team Digital Innovation - AI Center of Excellence. Quadro di lavoro orientato a progetti rivoluzionari, sviluppo e addestramento modelli ML/NLP/CV, e supporto a soluzioni basate su LLM e Computer Vision. Laurea STEM richiesta... 
    Consigliato

    PwC Italy

    Roma
    2 giorni fa
  •  ...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include...  ..., Larry Summers , and Jack Dorsey . Position: AI Safety Red Teamer Type: Contract Compensation: $70... 
    Consigliato
    Lavoro da remoto

    Mercor

    Roma
    un mese fa
  • Un importante gruppo industriale internazionale sta cercando Ricercatori/ Ricercatrici per i Sistemi Autonomi a Roma. La persona sarà coinvolta nello sviluppo di tecnologie per veicoli autonomi, migliorando il coordinamento tra agenti e implementando algoritmi innovativi...
    Consigliato
    Contratto con partita IVA
    Lavoro ibrido

    Leonardo

    Roma
    2 giorni fa
  • pHelloprima, a leading online motor insurance provider, is seeking a Senior Data Scientist to join our AI team in Rome. You will drive ML and Generative AI initiatives across operations such as claims, underwriting and customer support. /ppYou will collaborate with data... 

    Helloprima

    Roma
    1 giorno fa
  •  ...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include...  ..., Larry Summers , and Jack Dorsey . Position: AI Safety Practitioner Type: Contract Compensation: $... 
    Lavoro da remoto

    Mercor

    Roma
    un mese fa
  • 45.000 € - 65.000 €

     ...on a mission to empower businesses and professionals to leverage AI’s transformative capabilities. We believe technology should be...  ...the ever‑evolving AI landscape and collaborates with prominent research centres and organisations worldwide, such as the European Space... 
    Stage/Tirocinio

    Pi School

    Roma
    2 giorni fa
  • 48.000 € - 68.000 €

     ...has been awarded funding for the DVPS project, a collaborative research initiative worth 29 million euros over 4 years, which will be led...  ...of a team of researchers dedicated to DVPS within Translated's AI Research team, which works on several pieces of technology such... 
    Stage/Tirocinio

    Translated

    Roma
    2 giorni fa
  •  ...We are seeking experienced AI Safety Practitioners to evaluate the safety, quality, and alignment of frontier AI models across complex...  ...alignment and safety performance. Collaborate with AI researchers and safety teams on ongoing evaluation initiatives. Required... 

    Obsidian

    Roma
    2 giorni fa
  • ppWe are seeking experienced bAI Safety Practitioners /b to evaluate the safety, quality, and alignment of frontier AI models across complex, policy-sensitive, and ambiguous...  ...performance. /p /lilipCollaborate with AI researchers and safety teams on ongoing evaluation initiatives... 

    Obsidian

    Roma
    2 giorni fa
  • 50.000 € - 78.500 €

     ...enterprise clients, combining custom models with cloud and third-party AI services to deliver production-ready outcomes. Your role spans...  ...requirements - including edge device and HPC Engage in research and development of new AI and high-performance compute algorithms... 

    Accenture Italia

    Roma
    2 giorni fa
  •  ...Overview As an AI/ML Computational Scientist at Accenture, you will design, build, and operationalize AI solutions for enterprise clients across the full lifecycle. You’ll tailor models, including Deep Learning and Generative AI, and architect scalable MLOps pipelines... 

    Accenture

    Roma
    2 giorni fa
  •  ...Mercor is seeking experienced AI Safety Practitioners to evaluate safety, quality, and alignment of frontier AI models across complex...  ...Role focuses on identifying unsafe outputs, engaging with researchers and safety teams, and contributing to ongoing evaluation initiatives... 

    Mercor

    Roma
    2 giorni fa
  • 35.000 €

     ...Operationbr/Activities: computer vision, industrial automation, AI-powered quality control systemsbr/Techonologies: Casual/Generative...  ...inspection and environmental monitoring, taking solutions from research concept to deployed product. /p h3Responsibilities /h3 ul... 

    trtwo

    Roma
    2 giorni fa
  •  ...manifattura e operations, ai capital projects - per aiutare le aziende a reinventare l'intera catena del valore, alimentata da dati, AI e tecnologie di nuova generazione. /ppIn particolare, lavorerai su progetti di implementazione tecnologica e trasformazione digitale... 

    Accenture Italia

    Roma
    2 giorni fa
  • 35.000 €

     ...Operation Activities: computer vision, industrial automation, AI-powered quality control systems Techonologies: Casual/...  ...inspection and environmental monitoring, taking solutions from research concept to deployed product. Responsibilities Design, develop... 

    trtwo

    Roma
    2 giorni fa
  • 30.814,29 € - 35.000 €

     ...Data Artificial Intelligence, la quale dispone di tecnologie proprietarie, soluzioni e servizi che concretizzano il potenziale dell’AI e dei dati nell’evoluzione digitale di aziende e pubbliche amministrazioni. Può contare su oltre 450 clienti nazionali ed internazionali... 
    Tempo pieno
    Impiego permanente
    Lavoro ibrido
    Orario flessibile
    Dal lunedì al venerdì

    Almawave Labs

    Roma
    1 giorno fa
  •  ...Design Italy per definire requisiti EO/IR e HW per sistemi di bordo. Progetti, implementi e verifichi algoritmi di image processing e AI/ML per tracking, detection e riconoscimento. Collabori a soluzioni all’avanguardia in un contesto mission-critical, con attenzione a... 

    MBDA

    Roma
    2 giorni fa
  •  ...Computer Vision Engineer a Roma, in modalità ibrida. Collaborerai con il team di Product Engineering per sviluppare e integrare soluzioni AI e Computer Vision in contesti reali. Si richiede laurea magistrale o PhD in CV o discipline affini, 2–3 anni di esperienza,... 
    Lavoro ibrido

    Almawave Labs

    Roma
    2 giorni fa
  • 20.000 €

     ...Operation Activities: computer vision, industrial automation, AI-powered quality control systems Techonologies: Casual/...  ...solutions. An ideal role to grow skills in AI, imaging and applied research within EU projects. You will work under the guidance of senior... 

    trtwo

    Roma
    2 giorni fa
  • pAlmawave Labs Srl cerca un Junior Computer Vision Engineer da inserire nei Laboratori di Ricerca e Sviluppo Cognitive AI di Roma. L ruolo prevede affiancamento e training on the job, inizialmente con senior e gradualità fino all autonomia. /ppIn collaborazione con il... 
    Tempo pieno
    Impiego permanente
    Lavoro ibrido

    Almawave Labs

    Roma
    1 giorno fa
  • pDXC Technology Inc. is seeking a Data Scientist Consultant to support AI and Generative AI projects in Rome. You will design and implement AI solutions, collaborating with senior team members to translate client requirements into practical AI applications. /ppYou will... 
    Lavoro ibrido

    DXC Technology Inc.

    Roma
    2 giorni fa
  • 35.000 € - 45.000 €

     ...Data Artificial Intelligence, la quale dispone di tecnologie proprietarie, soluzioni e servizi che concretizzano il potenziale dell’AI e dei dati nell’evoluzione digitale di aziende e pubbliche amministrazioni. Può contare su oltre 450 clienti nazionali ed internazionali... 
    Tempo pieno
    Impiego permanente
    Lavoro ibrido
    Orario flessibile
    Dal lunedì al venerdì

    Almawave Labs

    Roma
    1 giorno fa
  •  ...Role: AI Data Scientist (Remote) Location: Remote (Work from Anywhere) Role Overview: We are hiring for one of our clients, seeking an AI Data Science Domain Expert to work on a Full-time basis. The role involves developing and optimizing AI models tailored to... 
    Tempo pieno
    Remoto

    Hire Feed

    Roma
    2 giorni fa
  • 29.000 € - 38.000 €

    pSpindox ricerca un AI Data Intelligence Research Engineer da inserire in team multidisciplinari per analizzare e valorizzare i dati nell'ambito Data Analytics, AI e ML. /ppSi richiedono PhD o laurea magistrale in campi affini, eccellente Python, conoscenza di PyTorch/TensorFlow... 
    Impiego permanente

    Spindox

    Roma
    1 giorno fa
  • Accenture Italia cerca uno Specialista di Computer Vision per il settore manifatturiero, all'interno di Supply Chain & Engineering. Progetterai soluzioni di CV basate su IA per contesti industriali, con focus su ispezione qualità, rilevamento difetti e ottimizzazione dei...

    Accenture Italia

    Roma
    2 giorni fa
  •  ...Salesforce is hiring Deployment Strategists — Strategic Builders who translate complex business challenges into agentic AI deployments that live and deliver value. You’ll own multi-faceted engagements, shape AI strategy, and ensure measurable outcomes for enterprise customers... 
    Tempo pieno

    Salesforce

    Roma
    2 giorni fa
  •  ...Deployment Strategist for the Public Sector, Nonprofit Education segments in Italy. You will own end-to-end customer engagements, define AI strategy, and ensure measurable adoption of agentic AI deployments across enterprise accounts. /ppThe role is client-facing with... 

    Salesforce, Inc.

    Roma
    2 giorni fa
  • pUniversità degli Studi di Roma Tor Vergata cerca un ricercatore con contratto a tempo determinato in regime di tenure track, presso il Dipartimento di Ingegneria Civile e Ingegneria Informatica, settore IINF-05/A. La sede di afferenza è nel Dipartimento di Ingegneria Civile...
    Tempo determinato

    Uniroma2

    Roma
    1 giorno fa