Open to AI Engineering roles

Hi, I'm Muhammad Ishaq

> ▍

I build intelligent systems powered by deep learning, language understanding and Generative AI — from reliable RAG pipelines and knowledge graphs to multilingual, voice-enabled assistants shipped to production.

Portrait of Muhammad Ishaq
RAG · LangChain
AI Agents · AWS
YOLO26 · ViT
85%faster retrieval
01 // about

From idea to launch, end to end.

I'm an AI Engineer based in Islamabad, Pakistan, focused on building smart systems that actually make it to production. My work spans Retrieval-Augmented Generation, knowledge graphs, conversational AI with multilingual and voice capabilities, and computer vision pipelines.

As Co-Founder of AI Cortexo, I lead the technical vision for a platform that lets businesses launch AI agents, workflows and RAG systems with minimal engineering effort.

Before that I led an AI & Data Science team, mentored junior engineers, and designed cloud-native ML architectures on AWS — Bedrock, SageMaker, Glue, ECS and EKS — always with an eye on latency, reliability and cost.

ishaq@ai: ~
$ whoami
co_founder · ai_engineer · team_lead · MS CS
$ cat focus.txt
[ "GenAI", "RAG", "Knowledge Graphs", "NLP", "Computer Vision" ]
$ languages --spoken
English · Urdu · Pashto
0Reduction in AI retrieval latency with knowledge graphs
0Uptime for production RAG pipelines
0Operational cost cut via inference optimisation
0Years shipping production AI systems
02 // experience

Where I've built things.

Co-Founder

AI Cortexo ↗ · Islamabad Hybrid Part-time

Jun 2025 — Present

Leading the technical vision and product development for an AI platform that lets businesses build, deploy and scale AI-powered applications with minimal engineering effort — enterprise-grade infrastructure for AI agents, workflow automation, RAG and intelligent business operations.

  • Architected and built the AI Cortexo Platform (AIP), enabling organisations to launch AI agents, AI workflows and intelligent automations from one unified platform.
  • Designed multi-agent orchestration systems supporting autonomous task execution and enterprise workflows.
  • Built scalable RAG pipelines powering enterprise document intelligence.
  • Developed production-ready AI infrastructure with Python, FastAPI, Docker, PostgreSQL, Redis and AWS.
  • Integrated OpenAI, Anthropic and open-source LLMs into enterprise workflows with intelligent model routing and cost optimisation.
  • Implemented authentication, usage tracking, API management and scalable cloud deployment pipelines.
  • Led product architecture, engineering strategy and AI research from concept to production, working directly with clients to automate business processes.
  • Built and launched Wink, a local-first AI desktop agent for Windows (Electron, TypeScript): it sees the screen with a local vision model via Ollama, then either guides the user click by click or operates the PC itself, with on-device Whisper speech-to-text and Kokoro voice, so screenshots never leave the machine.
AI AgentsMulti-Agent OrchestrationRAGFastAPIPostgreSQLRedisDockerAWSLLM RoutingElectronOllama
Visit aicortexo.com ↗ Try Wink ↗

Full-Stack AI Engineer

FlashSync AI ↗ · Lahore Remote Contract

Sep 2026 — Present

Building FlashSync AI, a pre-launch AI career-agent startup: first the production website and waitlist (from 4 design concepts to a live product on Vercel in about one week), now the product itself, the FlashSync Career Vault AI career agent.

  • Built the Career Vault AI career agent (Next.js 15, React 19, TypeScript): a zero-knowledge, local-first vault where users keep their profile, roles, wins and saved jobs, encrypted in the browser before anything is stored.
  • Designed envelope encryption with the Web Crypto API only: PBKDF2-SHA256 (600k iterations) derives a key that wraps a random AES-256-GCM data key; each record is sealed with AAD bound to its id and type, so tampered or swapped rows are rejected.
  • Shipped passphrase auth with no server accounts (unlock, 15-minute idle auto-lock, passphrase change by re-wrapping the key) and encrypted backup export/import across devices on Dexie.js/IndexedDB.
  • Added an AI STAR impact rewriter using Claude on Amazon Bedrock (EU region) that sends only the one bullet the user picks, plus a Manifest V3 Chrome extension that captures job posts from schema.org, Greenhouse and Lever pages.
  • Developed Sync, a Gemini RAG chatbot with streamed answers, embedding-based guardrails that keep it on-topic, English and Urdu replies, and voice input and read-aloud.
  • Engineered a custom viral waitlist with no third-party tools: referral-based queue ranking, early-access tiers, welcome emails, reCAPTCHA v3 and database-backed rate limiting on serverless functions with Turso.
  • Reached PageSpeed scores of 97–99 on mobile and 99–100 on desktop, cutting mobile LCP from up to 2.9 s to under 1.9 s with prerendered SEO pages; 0 WCAG 2.1 AA violations.
  • Delivered the QA suite used for client acceptance: Playwright across 6 browsers and devices, visual regression, a load test at 200 concurrent users with 0 errors, and 53 unit tests.
Next.jsReactTypeScriptWeb CryptoIndexedDBAWS BedrockChrome ExtensionGemini RAGTursoVercelPlaywright
Try the Career Agent (demo access on request) ↗ Visit flashsync.ai ↗

AI Engineer

Ventive LLC · Boise, ID Remote

Nov 2025 — Mar 2026
  • Built and deployed scalable AI and data systems, including NLP-powered intelligent chatbots and automated data pipelines on AWS.
  • Designed cloud data architectures with AWS Glue and Bedrock, using Docker-based deployments for production.
  • Developed and optimised ETL workflows to process large-scale enterprise data for analytics and AI use cases.
  • Engineered knowledge graphs to model complex entity relationships, reducing AI retrieval latency by up to 85% while maintaining high accuracy.
AWS BedrockAWS GlueKnowledge GraphsDockerETL

Team Lead — AI & Data Science

Octans Digital · Islamabad

Sep 2024 — May 2025
  • Led the AI/Data Science team driving development of scalable Generative AI and SaaS applications.
  • Built and deployed RAG pipelines with LangChain and AWS Bedrock, achieving 99.9% uptime and sub-second retrieval latency.
  • Designed inference optimisation strategies that cut operational costs by 25% while improving throughput.
  • Mentored junior engineers and apprentices, establishing structured practices for GenAI SaaS products.
LangChainRAGAWS BedrockLeadershipSaaS

Associate AI Engineer

Octans Digital · Islamabad

Mar 2023 — Sep 2024
  • Designed and deployed Generative AI solutions using state-of-the-art LLMs and AWS SageMaker for fine-tuning and inference.
  • Built multilingual NLP systems and computer vision pipelines focused on automation, accuracy and scale.
  • Developed AI chatbots, speech-to-text pipelines and intelligent document processing systems for enterprise workflows.
  • Implemented end-to-end ML workflows — data processing, training, deployment and monitoring — on ECS and EKS.
SageMakerLLM Fine-tuningSpeech-to-TextECS / EKSCV
03 // projects

Selected work.

A few systems I've designed and shipped across research and industry.

Octans Digital

Enterprise RAG Platform

Production RAG pipelines on LangChain and AWS Bedrock serving GenAI SaaS products with 99.9% uptime and sub-second retrieval, plus inference optimisations saving 25% in cost.

LangChainAWS BedrockDocker
Octans Digital

Multilingual Voice Assistant

Conversational AI with multilingual understanding, speech-to-text pipelines and smooth voice output, built to automate enterprise support workflows.

Speech-to-TextNLPLLMsTTS
Octans Digital

Intelligent Document Processing

LLM-driven extraction and classification of enterprise documents into structured data, streamlining manual, error-prone back-office processes.

LLMsSageMakerNLP
Ventive LLC

Cloud ETL & Data Pipelines

Cloud-native data architecture with AWS Glue and Docker deployments, processing large-scale enterprise data for analytics and downstream AI.

AWS GlueDockerETLPython
BS · HITEC University

Vision-Based Attendance System

Object-detection powered attendance with real-time identification and automatic record generation, replacing manual marking.

Object DetectionComputer VisionReal-time CV
BS Final Year Project

Defense Inventory Management

Web-based system digitising data entry, tracking and retrieval for a defense-sector client — secure, structured and built to replace legacy manual workflows.

C#ASP.NETWeb Forms
04 // lab

Built after hours, in the open.

Personal projects from my GitHub. Most started as a question I wanted to answer, from uncertainty-aware dental X-ray detection to a hand-written coding agent.

MuhammadIshaq-AIgithub.com ↗
0original repos
0years building in public
0vision models
  • Python 44%
  • Jupyter 33%
  • TypeScript 11%
  • Kotlin · JS · CSS
~/dentex-tooth-detectionPython

DENTEX: Uncertainty-Aware Dental X-ray Diagnosis

Fine-tuned YOLO26 on panoramic X-rays to detect caries, deep caries, periapical lesions and impacted teeth. Most detectors report a single confidence score, so I added Platt calibration, conformal guarantees and MC-dropout uncertainty. Impacted teeth reach 0.90 AP50. It's served on Vercel with FastAPI and ONNX Runtime.

YOLO26Conformal PredictionMC-DropoutONNXFastAPI
~/Agentic-WorkspacePython

Agentic Workspace

A terminal coding agent for a single directory. It reads, searches and edits files and runs commands. The agent loop is written by hand on the Anthropic Messages API with no framework, so every tool schema and permission check is visible in the code.

Tool UseClaude APIAgent Loop
~/oral-health-ragTypeScript

Dental Patient Assistant (RAG)

A chatbot for a dental practice in Islamabad. It answers only from 245 clinic, NHS and WHO passages and cites every answer. When the sources don't cover a question, it says so.

Next.jsGemini EmbeddingsVercel
~/Object-Detection-ML-APKotlin

Vision Lens: On-Device Detection

An Android app that detects 80 object types live from the camera with EfficientDet-Lite on LiteRT. It runs fully offline, so no image leaves the phone.

TFLiteCameraXJetpack Compose
~/VLM-TransformersJupyter

Vision-Language Model from Scratch

An 86M-parameter Vision Transformer built layer by layer in PyTorch, covering patch embedding, the CLS token and multi-head attention. It includes a Gradio app for captioning, VQA and attention maps.

PyTorchViTVQA
~/Serverless-RAGPython

Serverless RAG on AWS

A PDF upload to S3 triggers a Lambda that chunks, embeds and indexes the file into Pinecone. Queries go to a self-hosted Ollama model on EC2, which keeps the documents private.

AWS LambdaPineconeOllama
~/Safe-City-Computer-VisionPython

Safe-City License Plate Recognition

YOLO11 plate detection with EasyOCR, plus CLAHE enhancement for low light, glare, rain and motion blur. It outputs annotated traffic video.

YOLO11EasyOCROpenCV
~/Uni-Search-LLM-AgentPython

Scholarship Hunter Agent

An agent that finds fully funded Master's and PhD scholarships in Europe. It combines Gemini with Google Search grounding and portal scraping, removes duplicates in SQLite, and sends a notification for each new result.

GeminiStreamlitSQLite
~/Posture-Detection-ComputerVisionPython

Posture Guardian

Watches your webcam with MediaPipe Pose while you work. It flags slouching and uneven shoulders on a live dashboard and plays an alert if bad posture continues.

MediaPipeOpenCVFlask
~/nhanes-oral-health-mlPython

Periodontitis Risk, Explained

Interpretable ML on 10,683 NHANES adults using XGBoost and ordinal models, with SHAP explaining which risk factors matter. The labels match published CDC prevalence to within 0.1%.

XGBoostSHAPEpidemiology
~/Multilingual-LLM-sTypeScript

Multilingual NIM Studio

A Next.js app for multimodal chat and image generation on NVIDIA NIM models. API keys stay on the server, Redis handles rate limiting, and it ships as a Docker image.

NVIDIA NIMNext.jsRedis
~/transcription-servicePython

Transcription Service

A FastAPI service that turns audio into timestamped segments. It supports async jobs for long files and can switch to faster-whisper with one environment variable.

FastAPIWhisperffmpeg
$ ls ~/more-experiments 11 repos
05 // stack

Tools of the trade.

✦

Generative AI

LLMsAI AgentsMulti-Agent OrchestrationRAGLangChainLLM RoutingPrompt EngineeringFine-tuningKnowledge GraphsChatbots
◎

ML & Vision

Deep LearningNLPYOLO26Vision TransformersCNNsSpeech-to-Text
☁

Cloud & MLOps

AWS BedrockSageMakerAWS GlueECSEKSDockerKubernetes
⌘

Languages & Data

PythonFastAPINode.jsC#PostgreSQLRedisETLPower BITableau
06 // education

Learning, always.

Apr 2022 — Aug 2024

Master of Computer Science

FAST University, Islamabad

Thesis: Fabric DefectAI — deep learning and computer vision for automated fabric fault detection with YOLO26 and Vision Transformers.

Sep 2017 — Jul 2021

Bachelor of Computer Science

HITEC University, Taxila

Final year project: defense-sector inventory automation system, plus an object-detection based attendance system.

Certifications

  • AWS Certified Cloud Practitioner
  • Generative AI with Large Language Models
  • AWS Bedrock · ChatGPT Prompt Engineering for Developers
  • IBM — Introduction to Containers with Docker & Kubernetes
  • IBM — Data Science Methodology & Tools for Data Science
07 // contact

Let's build something intelligent.

Have a GenAI product, a RAG system that needs to be faster, or a vision problem to solve? My inbox is open.

ishaqjan619@gmail.com