Profile photo of Muhammad Saad

Muhammad Saad

Master's Student in Electrical and Computer Engineering | University of Ottawa | AI & Computer Vision

Graduate student specializing in Computer Vision and AI integration into immersive 3D environments.

I am a Master’s student in Electrical and Computer Engineering with a concentration in Artificial Intelligence at the University of Ottawa. My research focuses on pure computer vision and the integration of AI into immersive 3D worlds, including the development of AI-powered avatars and interactive virtual environments.

Prior to my graduate studies, I completed my Bachelor’s degree in Software Engineering at Islamia College Peshawar in 2021, where I developed expertise in medical imaging and computer vision alongside a strong foundation in software development and engineering principles.

My research focuses on:

  • Computer Vision: Pure computer vision techniques and applications
  • Immersive 3D Environments: AI integration into virtual worlds
  • AI-Powered Avatars: Intelligent virtual agents and interactive experiences
  • Medical Imaging: Facial expression recognition for neurological patients

Co-op Intern

May 2026 – Aug 2026
Dell Technologies Ottawa, Canada

Research topics: Agent Benchmarking and Agentic Operating Systems for Efficient Agent Execution on AI PC.

  • Conducting research on agentic operating systems for AI PCs, focusing on how intelligent agents can coordinate tasks, tools, and user workflows in next-generation personal computing environments.
  • Developing a human-agentic research assistant capable of taking a user-defined topic, decomposing it into research tasks, gathering relevant information, and generating structured, high-quality reports.
  • Exploring multi-agent reasoning, task planning, tool use, and human-in-the-loop refinement to improve the reliability, usefulness, and transparency of AI-assisted research workflows.

Software Engineer — AI

Sep 2025 – Feb 2026
EPAM Systems (Client: HUMAIN) Dubai, UAE
  • Deployed and customized an Open edX-based academy platform for HUMAIN Academy, including Tutor-based deployment, MySQL configuration, OAuth2 SSO integration, branded themes, enrollment workflows, and production-ready environment setup.
  • Built autonomous AI agents for learning-roadmap generation, enabling users to define career or skill goals and receive structured, personalized learning paths with recommended topics, resources, and progression plans.
  • Benchmarked and evaluated open-source voice cloning and text-to-speech models, focusing on inference performance, speech quality, model usability, and integration feasibility for AI-driven debate and conversational systems.
  • Worked on an autonomous debate system where LLM agents generated arguments, simulated debates between multiple AI debaters, and evaluated debate quality using agent-based judging and scoring pipelines.
  • Developed AI-camera research workflows for classroom intelligence, focusing on student emotion recognition, attention detection, and visual analytics to support adaptive learning and classroom engagement monitoring.

Graduate Research Assistant

Jan 2023 – Aug 2025
Metaverse Center, Mohamed Bin Zayed University of Artificial Intelligence Abu Dhabi, UAE

Research topics: Digital twin, Metaverse, Violence Detection, LLMs for Interactive Avatars.

  • Designed and launched WudFlux, a customized virtual learning platform built on Hubs, featuring full-body avatars, real-time lip-syncing, and sentiment-aware interaction for immersive web-based learning environments.
  • Built dTalk, an AI-powered avatar system with expressive 3D animations and real-time lip-sync using Mixamo and Three.js, enabling LLM-driven, speech-based interaction (GitHub).
  • Built expressive 3D avatar animations using Mixamo and enabled real-time lip-sync functionality through Three.js (GitHub).
  • Created a React-based analytics dashboard for the Malaria No More (MnM) project, enabling live data visualization and streamlined decision-making.
  • Developed a multimodal real-time violence detection system using LSTM, GRU, and Vision Transformer architectures on Jetson Nano at the Technology Innovation Institute (TII), improving inference performance for edge deployment.

Undergraduate Research Assistant

Dec 2020 – 2022
Digital Image Processing (DIP) Lab Islamia College Peshawar Peshawar, Pakistan

Research topics: Medical Imaging, Activity recognition, Facial emotion recognition (FER).

  • Contributed to NTNU’s implementation of the facial emotional recognition module assigned by the ALAMEDA AI Toolkit to analyze facial expressions for pain assessment and emotional state monitoring in neurological healthcare.
  • Attention-Based CNN-LSTM, CNN-GRU, and Video Vision Transformer (ViViT) Models for Complex Activity Recognition in Cricket.
  • Teaching assistant for Python programming course.

M.Eng. in Electrical and Computer Engineering (AI concentration)

2025–Present
University of Ottawa Ottawa, Canada

B.S. in Software Engineering

2017–2021
Islamia College Peshawar Peshawar, Pakistan

Multimodal Interaction through Embodied Agents: An Intelligent Assistant for Immersive Virtual Environments

M. Saad, M. Saeed, F. Laamarti, and A. El Saddik

1st ACM CHI 2026 Workshop on Shaping Future Human Connection: Social Augmentation through XR Technologies, Barcelona, Spain, 2026 — Accepted

CP-Diffusion: Conditional Prompt-Based Diffusion Models for Video Generation

M. Saeed, M. Khan, M. Saad, N. Rahim, W. Gueaieb, A. El Saddik

ACM Transactions on Multimedia Computing, Communications, and Applications (TOMM), 2025 — Accepted

Programming Languages

Python, MATLAB, C++, Java, SQL, JavaScript, TypeScript, HTML/CSS

AI & Machine Learning

PyTorch, TensorFlow, Keras, Scikit-learn, Hugging Face, OpenCV, Pandas, NumPy, Matplotlib, Vision Transformers, Diffusion Models, YOLOv5, CNNs, LSTMs, GRUs

LLMs & Agentic AI

LLaMA, DeepSeek, GPT-4, Claude, Groq, Mistral, LangChain, LangGraph, LlamaIndex, Agentic Workflows, Multi-Agent Systems, Prompt Engineering, LoRA/PEFT Fine-Tuning

RAG & Vector Databases

FAISS, Pinecone, ChromaDB, Qdrant, Semantic Chunking, Hybrid Retrieval, Embedding-Based Search

Backend & APIs

FastAPI, Flask, Node.js, REST APIs, WebSockets, JSON-RPC 2.0, OAuth2

MLOps & DevOps

Docker, Docker Compose, Git, GitHub Actions, CI/CD, Webhooks, Nginx, Caddy, Weights & Biases

Cloud & Databases

AWS S3, AWS EC2, AWS SageMaker, DigitalOcean, PostgreSQL, MySQL, MariaDB, Redis, Supabase

Platforms & Web/3D

Open edX, Tutor, Frappé LMS, Mozilla Hubs, React, Next.js, Three.js, A-Frame, Streamlit, Blender, WordPress, LaTeX

WudFlux

A web-based metaverse application featuring AI-powered professor avatars. Users can interact through audio or text, receiving natural responses to questions and queries in an immersive 3D virtual environment accessible through any browser.

DebateVerse

A fully autonomous debate system where two AI avatars engage in structured debates on given topics. Features intelligent moderator guidance, reference-supported arguments, and ensures proper engagement between debaters in real-time discussions.

Real-Time Violence Detection

Real-time violence detection system optimized for edge devices. Collaboration between Mohamed bin Zayed University of Artificial Intelligence (MBZUAI) and Technology Innovation Institute (TII). Deployed on Jetson Nano with model optimization techniques for efficient inference and resource-constrained deployment in real-world surveillance applications.

EmotionVis

Facial Expression Recognition system for monitoring neurological patients like Parkinson’s disease. Worked on training adaptive models for pain assessment and emotion monitoring under supervision of Dr. Sajjad at DIP Lab, Islamia College Peshawar. NTNU (Norwegian University of Science and Technology) collaboration funded by Alameda AI toolkit.