Open to PhD opportunities

Khalequzzaman Likhon

Machine learning engineer building reliable vision, video, and language systems.

ML Engineer, Accelx Inc. B.Sc. CSE, AUST Dhaka, Bangladesh
Portrait of Khalequzzaman Likhon

Research interests

What I work on

Across all of these, I care most about reliability: how models behave under real-world distribution shift, when they should abstain, and whether their outputs stay grounded in evidence that can be checked.

Questions I'm pursuing

02 · Real-Time Video Understanding

For detecting fights in surveillance video, which works better: appearance-based spatiotemporal features or explicit human-pose dynamics?

See: RWF-2000 comparison, Sentinel

03 · Large Language Models

How can small local models (3B-parameter LLMs, 2B VLMs) be made trustworthy enough for high-stakes decision support?

See: DocuRoute, SentinelRAG

04 · Generative AI

How do we keep generated output faithful to its source, whether that's generated code to its tests or a dubbed voice to its speaker?

See: Research→Code, voice-dub

About

A short introduction

I'm a Machine Learning Engineer at Accelx Inc., where I've built safety-critical perception and language systems since January 2025. These include real-time weapon, violence, and fall detection on live multi-camera video, and risk-assessment systems that write grounded, cited alerts. Watching these systems fail in the real world is why I focus on reliability.

My research asks whether strong results mean what they seem to. I've compared pose-based and appearance-based models for violence recognition, built LLM pipelines where every citation is checked, and written negative-result studies that test models against their shortcuts. One manuscript is under review at Applied AI Letters.

I hold a B.Sc. in Computer Science and Engineering from Ahsanullah University of Science and Technology. I'm now looking for PhD opportunities.

News

Recent updates

Selected projects

Things I've built

All projects →

Let's talk

I'd love to hear from faculty recruiting PhD students and from researchers working on vision, video, or language models.