DINOv3-Mask2Former: Urban Instance Segmentation
Unified instance segmentation framework combining self-supervised DINOv3 Vision Transformer with Mask2Former for urban streetscapes
Every project on this site: 25 of them, from instance segmentation for urban streetscapes to a firefighting robot and a CNC machine built from scratch. Cards with a title link lead to a full write-up.
Unified instance segmentation framework combining self-supervised DINOv3 Vision Transformer with Mask2Former for urban streetscapes
High-performance Rust library for rendering rich text and complex terminal user interfaces
Cross-lingual learning model for effective Arabic text-to-image retrieval using knowledge distillation from CLIP
Interactive visualizer for the Mapillary Vistas semantic segmentation dataset
Implementation of Structure from Motion algorithm for 3D reconstruction from multiple images
Text-to-speech system specialized for Arabic poetry with proper pronunciation and rhythm
Large-scale Arabic speech corpus for automatic speech recognition research
Integration of RASA framework with ChatGPT for enhanced conversational AI
Comprehensive RGB dataset for Arabic Alphabets Sign Language recognition
Real-time Arabic sign language recognition system deployed on Raspberry Pi
Deep learning models for Arabic Sign Language recognition
Generative Adversarial Networks implementation for MNIST digit generation
Neural network for MNIST handwritten digit classification
Recurrent Neural Network for sentiment analysis of text data
Implementation of Word2Vec algorithm using skip-gram architecture