Skip to content
View Manishmaurya89's full-sized avatar

Block or report Manishmaurya89

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Manishmaurya89/README.md
███╗   ███╗ █████╗ ███╗   ██╗██╗███████╗██╗  ██╗
████╗ ████║██╔══██╗████╗  ██║██║██╔════╝██║  ██║
██╔████╔██║███████║██╔██╗ ██║██║███████╗███████║
██║╚██╔╝██║██╔══██║██║╚██╗██║██║╚════██║██╔══██║
██║ ╚═╝ ██║██║  ██║██║ ╚████║██║███████║██║  ██║
╚═╝     ╚═╝╚═╝  ╚═╝╚═╝  ╚═══╝╚═╝╚══════╝╚═╝  ╚═╝
                                     Manish Maurya

B.Tech CSE · AI / ML · Computer Vision · NLP

Building intelligent systems — from real-time video analytics to multilingual data tools


About Me

B.Tech Computer Science graduate from REVA Institute of Technology, Bangalore (2021–2025), specialising in AI and Machine Learning.

I build practical ML systems — real-time computer vision pipelines, NLP data tooling, and emotion-aware applications. My current focus is on deep learning, multilingual NLP, and making AI work for underrepresented languages and low-resource settings.

profile = {
    "name":       "Manish Maurya",
    "degree":     "B.Tech CSE — REVA Institute of Technology, Bangalore (2021–2025)",
    "focus":      ["Deep Learning", "Computer Vision", "NLP", "Multilingual AI"],
    "stack":      ["Python", "PyTorch", "TensorFlow", "OpenCV", "Streamlit"],
    "email":      "manish.maurya0408@gmail.com",
    "location":   "India "
}

🚀 Featured Projects

Project What It Does Tech Stack
Basketball Analytics Detects & tracks players, ball, and court elements from game footage. 95%+ detection accuracy, real-time heatmaps & team stats YOLOv8 · DeepSORT · OpenCV · PyTorch · Docker
Music × Emotion Webcam-based emotion detection recommends music in real time. CNN inference optimised by 25% Python · CNN · OpenCV · TF/Keras · Streamlit
Face Direction Detection Real-time head orientation system with Kalman filtering. Improved detection accuracy by 30% Python · CNN · OpenCV · Kalman Filter
Synthetic Data Generator Generates multilingual NLP datasets in 13 languages (Hindi, Urdu, Tamil, Telugu + more) via local LLM Python · Ollama · ReportLab · gTTS

🛠️ Skills

Languages & Frameworks Python TensorFlow PyTorch Scikit-learn OpenCV YOLOv8

Data & Visualisation Pandas NumPy Matplotlib

Tools & Platforms Docker Git Jupyter Notebook VS Code Streamlit SQL


🏅 Certifications

  • 🎓 Machine Learning Specialization — DeepLearning.AI
  • 🎓 TensorFlow Developer Professional Certificate — DeepLearning.AI
  • 🎓 Python for Everybody — Udemy
  • 🎓 Python for Data Science — IBM SkillsBuild
  • 🎓 HTML & CSS — Coursera / Johns Hopkins University


🧠 Engineering Philosophy

  • If it’s not measurable, it’s not improving

  • If it’s not reproducible, it’s not engineering

  • If it doesn’t scale, it’s a prototype

  • Building myself


📧 manish.maurya0408@gmail.com  |  📍 India

""I don’t just build models — I build systems that work in the real world."*

Pinned Loading

  1. Basketball-Analytics-Estimation-Using-Object-Segmentation- Basketball-Analytics-Estimation-Using-Object-Segmentation- Public

    Basketball game analysis tool using deep learning for object detection, player tracking, and team assignment.

    Jupyter Notebook 1

  2. Synthetic-data-generator- Synthetic-data-generator- Public

    Multilingual synthetic text dataset generator powered by Ollama. Supports 13 South Asian languages with TXT, JSON, CSV, PDF and audio output.

    Python

  3. Face-direction-detection Face-direction-detection Public

    PROJECT

    JavaScript

  4. DeepSpeech DeepSpeech Public

    Forked from mozilla/DeepSpeech

    DeepSpeech is an open source embedded (offline, on-device) speech-to-text engine which can run in real time on devices ranging from a Raspberry Pi 4 to high power GPU servers.

    C++

  5. GLiNER GLiNER Public

    Forked from urchade/GLiNER

    Generalist and Lightweight Model for Named Entity Recognition (Extract any entity types from texts) @ NAACL 2024

    Python

  6. MUSIC-RECOMMENDATION-USING-FACIAL-RECOGNITION MUSIC-RECOMMENDATION-USING-FACIAL-RECOGNITION Public

    Music Recommendation using Facial Expression

    Python