AI / ML Engineer

Osama Sleem

Computer Vision & Generative AI — Cairo, Egypt

I build production computer-vision and generative-AI systems — from diffusion-based try-on and 3D point-cloud pipelines to bilingual speech recognition — and ship them with the MLOps to back it up.

// snapshot

  • ASR infra cost reduction80–90%
  • Manual aligner-prep time cut60%
  • 3D staging directional F10.99
  • Shipped projects & models9+
  • Peer-reviewed publication1

Featured Work

Three systems, start to finish

ZAY Marketplace virtual try-on interface showing a generated outfit on a model with a fit-rating prompt
Generative AIProduction

ZAY Marketplace — AI Platform

Diffusion-based virtual try-on for realistic garment transfer, plus a visual search & product-similarity engine on deep embeddings and FAISS. Deployed as scalable FastAPI/Docker inference services with async processing and GPU acceleration, backed by a full MLOps pipeline.

Diffusion modelsFAISSFastAPI + DockerMLOps
3D Computer VisionMedical

3D Aligner Orthodontics Planning

End-to-end pipeline predicting cumulative and stage-wise tooth transformations. Built and annotated a 3D dental dataset via manual segmentation, then used DGCNN and PointNet++ for point-cloud feature extraction with a Transformer decoder to model sequential jaw movements.

−60% manual prep timestage loss 0.09directional F1 0.99
Medical ImagingClassification

X-ray AI Diagnostic Assistant

TensorFlow classifier for Tuberculosis, Pneumonia and Knee Osteoarthritis using a custom CNN, benchmarked against fine-tuned VGG and ResNet backbones. Integrated LangChain to power a chatbot that interprets diagnoses and communicates results interactively to patients.

Custom CNN F1 92%VGG F1 94%ResNet F1 89%

More Projects

Research, experiments & competitions

Facial attribute manipulation results showing input, inverted, and edited output faces
Generative AIVision

Facial Attribute Manipulation — CLIPInverter + PTI

A CLIPInverter + Pivotal Tuning Inversion pipeline for text-driven facial edits with high realism and identity retention.

L2 0.066LPIPS 0.243ID loss 0.432
View on GitHub ↗
Text-to-image interface generating an X-ray image from a symptom description prompt
Generative AIMedical Imaging

Symptom-to-Medical-Image Generator

Fine-tuned Stable Diffusion with LoRA to generate X-ray, CT and MRI images directly from natural-language symptom descriptions.

train loss 0.09LoRA fine-tune
View on Hugging Face ↗
AI-generated modern home office interior design from a text prompt
Generative AIDesign

Room Design Image Generation

Customized Stable Diffusion via Hugging Face Diffusers to generate realistic room designs from natural-language prompts, trained on 1,000+ paired room images.

+30% style variation1,000+ image dataset
View on Hugging Face ↗
Super-resolution results showing a pixelated input face restored to a sharp output image
Computer VisionGANs

Face Photo Resolution (SRGAN)

A Super-Resolution GAN built from scratch to upscale low-resolution facial images, validated with consistent qualitative gains in clarity and sharpness.

SSIM 0.74PSNR 23.39 dB
View on Kaggle ↗
⚽No public demo
Machine LearningSports Analytics

Football Player Analysis Modeling

ML models evaluating player rating, market value and potential to support talent scouting and recruitment planning.

Rating 97%Market value 99%Potential 94%
View on Kaggle ↗
🌱No public demo
Machine LearningAgriTech

Turba — AI-Powered Smart Farming System

Crop recommendation model trained on soil datasets to improve reliability for small and mid-scale farms. 1st place, Samsung climate hackathon.

99% classification accuracy
Ask about this project →

About

Machine learning across domains

Machine learning engineer skilled in applying AI across diverse domains such as medical imaging and real estate design. I've led a 3D aligner orthodontics planning pipeline, optimized X-ray imaging for diagnostics, and fine-tuned generative models for real estate visualization — always with an eye toward training and deploying models that hold up in production.

EXPERIENCE
Mar 2026 — Present
Co-Founder & AI Engineer
ZAY Marketplace · Part-time, Hybrid
  • Built the virtual try-on and visual search/FAISS retrieval systems, deployed via FastAPI/Docker with async, GPU-accelerated inference.
  • Owns the MLOps pipeline: containerization, deployment, versioning, monitoring, CI/CD.
Nov 2025 — Mar 2026
AI Engineer
DITC · Remote
  • Built bilingual (Arabic–English) ASR models optimized for CPU-only inference at ~25% WER, cutting infra cost 80–90% vs. GPU deployment.
  • Owned the full ASR pipeline: audio preprocessing, custom tokenization, training, evaluation, production integration.
Sep 2025 — Present
Computer Vision Instructor
AI Coding Academy · Remote
  • Teach Deep Learning, Computer Vision, and 3D Modeling; mentor students shipping real projects with PyTorch, TensorFlow, and OpenCV.
Sep 2024 — Nov 2024
AI Intern
Dar Global · Remote, UAE
  • Fine-tuned Stable Diffusion models for automated real estate visualization.
  • Assessed AI tools across business units to guide technology adoption.
Jul 2023 — Nov 2023
AI Intern
Samsung Innovation Campus · Cairo, Egypt
  • Supervised/unsupervised learning, object detection, classification, and GANs — 2nd place among AI interns.
  • Ran end-to-end ML workflows and presented results internally.
EDUCATION
Master's, Artificial Intelligence
Cairo University · Feb 2026 — Present
Bachelor's, Artificial Intelligence
Kafr Elshiekh University · Jan 2021 — Jun 2025

GPA 3.7/4.0 · Top 5 Students, Artificial Intelligence college

SKILLS

Computer Vision & Generative AI

Image classificationObject detectionSegmentationObject trackingGANsDiffusion ModelsStable Diffusion

3D & Deep Learning

Mesh processingPoint clouds (DGCNN, PointNet++)STL manipulationCNNsLSTM / RNN

MLOps & Engineering

DockerFastAPIModel deployment & servingCI/CDREST APIsGPU inferenceModel versioningMonitoringLinuxGit

Libraries & Languages

PythonC++PyTorchTensorFlowOpenCVPandasNumPyScikit-learnDiffusersHugging Face
PUBLICATIONS & AWARDS
Evaluating Deep Learning Architectures for Multi-Class Classification of Medical X-ray Images
FICAC25, 2025

Comparative study of Custom CNN, VGG-16 and ResNet for diagnosing tuberculosis, pneumonia and osteoarthritis — up to 96% accuracy with VGG-16.

Top 5 Students
AI college, Kafr Elshiekh University
1st Place, Climate Change Hackathon (Turba) & Top 10, Unigreen Competition
Samsung Innovation Campus

Contact

Let's talk about computer vision, generative AI, or shipping ML to production

Open to roles and collaborations in computer vision, generative AI, and MLOps. Based in Cairo, Egypt — available remote or hybrid.