About Experience Publications Certificates News Projects

Oudoum Ali Houmed

Oudoum Ali Houmed

AI Safety & Security

Currently: Research Engineering Intern at Kappa Santé — Generative Modeling & AI Robustness · Healthcare

About

I'm an M.Sc. student in Data, Knowledge and Hybrid Artificial Intelligence (DKAI) at Paris-Saclay University and an AI Safety Research Fellow at Black in AI Safety & Ethics under Krystal Jackson (UC Berkeley, Center for Long-Term Cybersecurity). My fellowship project is “Can We Trust Deception Monitors for AI Agents? A Robustness-Gap Protocol and the Limits of Adaptive Evasion.”

My fellowship work asks how much deception detection survives as adversary budget rises: a robustness-gap protocol for monitors on AI agents (surface classifiers, residual-stream probes, and CoT controls), with a retention gate so a broken agent is not scored as evasion. Broader interests include adversarial robustness, red teaming, mechanistic interpretability, and control mechanisms for frontier models.

Alongside the fellowship, I work as a Research Engineering Intern at Kappa Santé on constraint-guided virtual patient generation for clinical datasets. Previously, I was a Research Assistant at ANÖROM, where I studied the adversarial robustness of medical diagnostic models. I have completed the ARENA curriculum, OxML 2025, and BlueDot Impact's Technical AI Safety course.

Experience

Kappa Santé

May 2026 – Sep 2026
Research Engineering Intern · Generative Modeling & AI Robustness · Healthcare · Paris, FR

Applied research on constraint-guided virtual patient generation for longitudinal clinical datasets: conditional generation pipelines with explicit filtering for statistical fidelity and temporal coherence, and evaluation of downstream model safety with a focus on calibration and generalization under distribution shift.

Black in AI Safety & Ethics

May 2026 – Jul 2026
AI Safety Research Fellow · Alignment & Security Track · Remote

Fellowship project under Krystal Jackson (UC Berkeley CLTC): Can We Trust Deception Monitors for AI Agents? A Robustness-Gap Protocol and the Limits of Adaptive Evasion — on Llama-3.3-70B with the published Apollo probe, an adversary-budget ladder (b0–b4), and a retention gate before reporting detection drop. Findings feed technical standards, safety controls, and policy recommendations. Completed the intensive 5-week ARENA curriculum.

ANÖROM — Big Data & AI Dept

Oct 2024 – Jun 2025
Research Assistant · AI Robustness · Ankara, TR

Investigated the vulnerability of medical diagnostic architectures to adversarial perturbations using gradient-based optimization. Implemented white-box attacks (FGSM, PGD) to quantify the robustness of tumor detection models; contributed empirical findings toward a forthcoming manuscript (Adversarial Threats to Safety-Critical Medical AI).

ADEO Cybersecurity

Feb 2024 – Jun 2024
Cybersecurity Specialist Intern · Istanbul, TR

Achieved an internship grade of 95%. Detected and analyzed RDP brute-force attacks using Wazuh SIEM; performed deep-packet inspection with Wireshark and web-app pen-testing with Burp Suite; conducted malware scanning, static/dynamic reverse engineering, and forensic analysis with MDR tools.

AI Security & Defence Lab (AISEC LAB)

Jul 2023 – Sep 2023
Summer Research Intern · Konya, TR

Co-developed Cyber Inspector, a WAF-integrated ML system for malicious query detection. Managed the full project lifecycle: data collection, preprocessing, model training (SVM, RF), and deployment (Detection of Malicious Web Queries with Machine Learning).

Education

M.Sc. Data, Knowledge and Hybrid Artificial Intelligence (DKAI) — Paris-Saclay University

Paris, FR. Relevant coursework: Towards Adaptive and Agentic AI; Exploring Data Declaratively.

Ranked #1 in Europe in Mathematics (Shanghai Ranking).

Sep 2025 – present
B.Sc. Computer Engineering — Necmettin Erbakan University

Top 10% of cohort. Full scholarship. Konya, TR.

2020 – 2024

Research Stack

Libraries & Frameworks

PyTorchJAXTransformerLens HuggingFaceWandBTracr

Core Concepts

Adversarial MLMech InterpRLHF Activation SteeringProbe Training Red TeamingThreat Modeling

General

PythonC++PySpark Linux/BashGitDocker LaTeX

Publications

Artificial Intelligence Studies paper: Endpoint Security review

Artificial Intelligence Studies · 2025

A Systematic Review on Evolution, Challenges, and Future Trajectories of Endpoint Security: Integrating Zero Trust and Federated Learning Perspectives

O. A. Houmed, O. Ceran

Cover of Adversarial Threats to Safety-Critical Medical AI manuscript

Manuscript · ANÖROM Big Data & AI · 2025

Adversarial Threats to Safety-Critical Medical AI: A Security Assessment of 20 Deep Learning Tumour Detectors in Brain MRI and Kidney CT

O. A. Houmed

In preparation · Target: AI4Good @ NeurIPS Workshop

Can We Trust Deception Monitors for AI Agents? A Robustness-Gap Protocol and the Limits of Adaptive Evasion

O. A. Houmed, K. Jackson

Cover of Detection of Malicious Web Queries with Machine Learning

AISEC LAB · 2023

Detection of Malicious Web Queries with Machine Learning

A. Karaca, Z. Güney, O. Kortun, O. A. Houmed

Certificates

CompTIA Security+ ce certificate

CompTIA · Oct 2023

CompTIA Security+ ce

Valid through Oct 2026 · Verify at verify.CompTIA.org

UNDP Frontier Tech Leaders Machine Learning Bootcamp certificate

UNDP ICPSD · SDG AI Lab · UN Technology Bank · 2023

Frontier Tech Leaders Programme — Machine Learning Bootcamp

Certificate of Completion

Highlights / News

2026
May

Selected for the AI Safety Research Fellowship at Black in AI Safety & Ethics (Alignment & Security track), supervised by Krystal Jackson (UC Berkeley CLTC). Started Can We Trust Deception Monitors for AI Agents?, measuring how monitors hold up under adaptive adversaries.

May

Joined Kappa Santé in Paris as a Research Engineering Intern, working on constraint-guided virtual patient generation.

 

Formalized the Narrative Susceptibility Score (NSS) in the LLM Biases for Political Conspiracies project, with a 600+ prompt evaluation protocol.

2025
Sep

Started the M.Sc. in Data, Knowledge and Hybrid Artificial Intelligence (DKAI) at Paris-Saclay University.

 

First paper published in Artificial Intelligence Studies: “A Systematic Review on Evolution of Endpoint Security” (with O. Ceran).

 

Attended OxML 2025 (Oxford Machine Learning Summer School) and completed BlueDot Impact's Technical AI Safety course.

2024
Oct

Joined ANÖROM (Big Data & AI Dept) in Ankara as a Research Assistant on AI robustness.

Feb

Cybersecurity Specialist Intern at ADEO Cybersecurity (Istanbul): Wazuh SIEM, Wireshark, Burp Suite, and MDR forensics (internship grade 95%).

 

Graduated in the top 10% of the B.Sc. Computer Engineering cohort at Necmettin Erbakan University, on a full scholarship.

2023
Jul

Summer Research Intern at AISEC LAB (Konya): co-developed Cyber Inspector, a WAF-integrated ML system for malicious query detection.

 

Completed the UNDP-IICPSD Machine Learning Bootcamp.

Projects