Skip to content
View 07anishu12's full-sized avatar

Highlights

  • Pro

Block or report 07anishu12

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
07anishu12/README.md

Aniket Thakur

Applied AI & Platform Engineer

I design and build reliable, production-grade AI applications, document intelligence pipelines, and data platform infrastructure. I bridge the gap between machine learning capabilities and scalable software engineering.


πŸ› οΈ Tech Stack & Tooling

  • Languages: Python, TypeScript, SQL (PostgreSQL, MySQL), Bash
  • AI & ML: LLM Prompt Engineering, RAG (pgvector, ChromaDB), PaddleOCR, Computer Vision (MediaPipe, OpenCV)
  • Backend & APIs: FastAPI, Node.js (Express), REST APIs, WebSockets
  • Frontend: Next.js (React), Zustand, TailwindCSS, Recharts
  • Infrastructure & DevOps: Docker, Docker Compose, Nginx, Turborepo, pnpm Workspaces, CI/CD (GitHub Actions)

πŸš€ Featured Projects

A natural-language business intelligence application that translates user queries into audited SQL and interactive charts.

  • Stack: Python, FastAPI, SQLAlchemy 2, Alembic, React, TypeScript, Zustand, Recharts, Docker.
  • Key Features: Implemented a model-backed prompt engine, safe SQL query validation schemas, database connector adapters (Postgres, DuckDB, MySQL), and a grid-based canvas.

An enterprise document processing and form automation platform with OCR and structured extraction.

  • Stack: Next.js (TypeScript), TailwindCSS, FastAPI, Python, PaddleOCR, Docker.
  • Key Features: Engineered a decoupled OCR microservice, structured layout extraction pipelines, and automated UI integration tests via Playwright.

Full-stack invoice extraction with local OCR, structured parsing, and deterministic offline fallbacks.

  • Stack: Python, FastAPI, React, TypeScript, PaddleOCR, SQLite, Pydantic.
  • Key Features: Combined OCR text extraction with Pydantic schema validation for deterministic extraction of items, vendor info, and total values.

An LLM-driven research, content-generation, and automated publishing system.

  • Stack: JavaScript, Node.js, Puppeteer, Python, Gemini API.
  • Key Features: Automated web scraping, dynamic infographic generation via HTML-to-image renderers, and automated multi-channel publishing scheduling.

Cache-first Node.js API for daily fuel-price data across Indian cities and states.

  • Stack: JavaScript, Node.js, Express, Redis.
  • Key Features: Implemented cache-first retrieval logic, automated daily scraping of state fuel prices, and structured REST API endpoints.

πŸ“¬ Connect with Me

Pinned Loading

  1. india-document-tools india-document-tools Public

    Document intelligence and form-processing platform for India with OCR, Next.js dashboard, and FastAPI structured extraction microservice.

    TypeScript

  2. invoice-ocr-extractor invoice-ocr-extractor Public

    Full-stack invoice extraction with PaddleOCR, structured parsing, and deterministic fallbacks.

    Python

  3. multi-platform-content-automation multi-platform-content-automation Public

    Autonomous LLM-powered content generation, image rendering, and browser-automation publishing pipeline.

    Python

  4. prompt-bi prompt-bi Public

    Natural-language BI platform turning prompts to validated SQL, charts, and interactive dashboards. Built with FastAPI, React, SQLAlchemy 2, and Docker.

    TypeScript

  5. Fitness-tracker- Fitness-tracker- Public archive

    This machine learning project analyzes data from wearable fitness devices, focusing on gyroscope and accelerometer readings to track and predict user activities and health metrics

    Python 1

  6. jobs-tracker jobs-tracker Public

    Python