Skip to content
View mfayazkhan50-AI's full-sized avatar

Block or report mfayazkhan50-AI

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
mfayazkhan50-AI/README.md

markdown

About Me

I am Muhammad Fayaz Khan, an LLM Engineer and ML practitioner from Pakistan. I first mastered data communication using Pandas and NumPy. Later, I specialized in predictive modeling and machine learning with Scikit-Learn then Deep artificial neural networks and deep learning with TensorFlow and Pyrorch. Currently, I focus on developing enterprise-grade RAG systems and Agentic AI workflows using FastAPI, LangChain, and Vector Databases. I focus primarily on maximizing retrieval accuracy, implementing latency optimization, and achieving LLM cost reduction for production-ready AI applications. As a self-taught professional, I am executing an advanced LLM Engineering roadmap in 2026.

Socials

Facebook Instagram LinkedIn TikTok email GitHub

Tech Stack

Python PyTorch TensorFlow Keras scikit-learn NumPy Pandas SciPy Matplotlib Plotly Streamlit FastAPI LangChain LangGraph CrewAI ChromeDB Qdrant Pinecone Docker Git GitHub AWS Google Cloud Azure OpenAI Anthropic Google Gemini Groq MongoDB MySQL SQLite C++ Cloudflare Heroku MicrosoftSQLServer Canva mlflow

Projects Portfolio

BizBrain Business PDF Agent

Enterprise FAQ agent designed to ingest business PDFs and answer domain-specific queries using Hybrid Search and Cross-Encoder Reranking workflows. The architecture is built using FastAPI, ChromaDB, LangChain, Pydantic, Chainlit, and Docker. It is engineered for real client deployment featuring persistent chat history and exhaustive RAGAS evaluation.

Enterprise RAG Platform

Production-grade RAG system utilizing Recursive Semantic Chunking and Cross-Encoder Reranking strategies to query private PDF data safely. The pipeline achieved high faithfulness and answer relevancy scores validated via the RAGAS evaluation framework. Technical stack includes FastAPI, ChromaDB, LangChain, and RAGAS.

Multi-Functional AI Agent

Autonomous agent built with structured Tool Calling mechanics and validation pipelines producing precise JSON outputs via Pydantic v2. Developed to handle complex, multi-step business workflows autonomously, engineered using LangChain, and optimized thoroughly for multi-agent cooperation systems.

Rossmann Sales Predictor

Predictive model engineered for the Rossmann retail chain to execute highly accurate sales forecasting tasks. Applied advanced feature engineering strategies along with ensemble methods to achieve competitive accuracy metrics on time-series retail datasets. Technical stack utilizes Python, Scikit-Learn, and Pandas.

Churn Predictor

Machine learning classification model constructed to predict customer churn trends. The codebase contains full end-to-end data preprocessing pipelines, strategic feature selection setups, and extensive model evaluation stages using precision-recall analysis. Technical stack includes Python, Scikit-Learn, Pandas, and FastAPI.

Diabetes Risk Analyzer

Classification model deployed successfully as an interactive Streamlit web application to provide instant diabetes risk prediction based on complex patient health parameters. Technical stack utilizes Python, Scikit-Learn, and Streamlit.

House Price Prediction

End-to-end regression pipeline developed for predicting real estate house pricing trends. The architecture incorporates advanced data cleaning, exploratory data analysis, and multi-model comparison matrices. Technical stack includes Python, Scikit-Learn, and Streamlit.

Cat vs Dog CNN Classifier

Deep learning image classification model utilizing Convolutional Neural Networks built entirely with PyTorch. The model is deployed as an interactive user interface via Streamlit. Technical stack utilizes PyTorch, custom CNN layers, and Streamlit.

Key Competencies

  • Optimizing LLM API costs and execution latency metrics for production-ready client applications
  • Implementing production LLM Caching mechanisms via GPTCache, reducing operational API costs by up to 40 percent
  • Engineering stateful, cyclical multi-agent orchestration architectures and workflows utilizing LangGraph
  • Designing robust RAG 2.0 pipelines featuring Hybrid Search architectures and Neural Reranking systems for enterprise environments
  • Interfacing and integrating multiple core LLM providers including OpenAI, Anthropic, Gemini, and Groq
  • Executing red-teaming methodologies and evaluating generative LLM outputs using RAGAS and LLM-as-Judge frameworks
  • Formulating advanced Prompt Engineering designs including Zero-shot, Few-shot, Chain-of-Thought, ReAct Patterns, and Context Engineering
  • Developing Agentic AI frameworks utilizing LangGraph, CrewAI, LangChain, Tool Calling methodologies, and Multi-Agent Workflows
  • Structuring backend services and APIs using FastAPI, asynchronous Python, decorators, generators, Pydantic v2, and the Instructor library
  • Managing vector databases effectively including production setups with ChromaDB, Qdrant, and Pinecone
  • Applying foundational AI and ML principles including Machine Learning, Deep Learning, PyTorch, CNN architectures, and Scikit-Learn pipelines
  • Implementing DevOps tooling setups including Docker, Docker Compose, Git, GitHub version control, and Streamlit deployment

Education

  • 12th Grade, Intermediate Computer Science, Completed
  • Self-Taught LLM Engineering Curriculum, 2026

Location

  • Lakki Marwat, KPK, Pakistan, also operating within the Rawalpindi and Islamabad region

Contact

Pinned Loading

  1. UrduGPT UrduGPT Public

    End-to-end LLM engineering project for Urdu using Supervised Fine-Tuning (SFT) and Direct Preference Optimization (DPO).

    Jupyter Notebook 1

  2. Python-Bootcamp-batch-1 Python-Bootcamp-batch-1 Public

    Python Course Coding Materials

    Python 1