STARTUP-AI-AGENT

agent
Security Audit
Warn
Health Warn
  • No license — Repository has no license file
  • Description — Repository has a description
  • Active repo — Last push 0 days ago
  • Low visibility — Only 5 GitHub stars
Code Pass
  • Code scan — Scanned 12 files during light audit, no dangerous patterns found
Permissions Pass
  • Permissions — No dangerous permissions requested

No AI report is available for this listing yet.

SUMMARY

AI-powered startup intelligence agent — hand-built ReAct loop, parallel tool execution, live web search, and RAG-grounded document analysis. No frameworks. Full understanding.

README.md
Typing SVG

🧠 BizRadar AI

AI-Powered Startup Intelligence & Business Analysis Agent


Python
Groq
Llama
Gemini
Tavily
ChromaDB



License
Version
Phase 3
Phase 4
No Frameworks



Transform any startup idea into a structured business intelligence report — powered by a hand-built ReAct agent, parallel tool execution, real-time web search, and RAG-grounded document analysis.


Getting StartedFeaturesArchitectureRoadmapContact


🎯 The Problem This Solves

You have a startup idea. You need answers — fast, cited, and trustworthy.

The problem with generic AI tools: they hallucinate. They produce confident market statistics that don't exist. They cite reports that were never written. You can't put that in an investor pitch deck.

BizRadar AI solves this with three layers of grounding:

  • Real-time web search via Tavily — every market claim has a source URL
  • Document-grounded RAG via ChromaDB — answers from your own pitch deck, not the LLM's memory
  • ReAct reasoning loop — the agent reasons about what it needs before acting, not after

Example: "AI-powered tiffin service for students"

Output: A complete startup analysis — market potential, competitors, MVP features, tech stack, risk assessment — every claim cited, every document answer grounded.


✨ Features

Feature Description Status
⚡ Groq LPU Inference Ultra-fast token generation via purpose-built LPUs ✅ Live
🔁 ReAct Agent Loop LLM reasons → calls tools → observes → loops to final answer ✅ Live
🧵 Parallel Tool Execution Tools run simultaneously via ThreadPoolExecutor ✅ Live
🧠 Sliding Window Memory Remembers last 6 conversation turns ✅ Live
🔍 Real-Time Web Search Live market analysis via Tavily API with cited sources ✅ Live
📊 Startup Analysis Structured business intelligence report ✅ Live
💡 MVP Recommendations Core feature suggestions via Gemini 2.5 Flash ✅ Live
🏗️ Tech Stack Advisor CTO-level stack recommendations via Gemini 2.5 Flash ✅ Live
⚠️ Risk Analysis Fatal flaw identification and mitigation strategies ✅ Live
📄 PDF Document Ingestion Ingest pitch decks and business documents ✅ Live
🗄️ Vector Search (RAG) Semantic search over ingested documents via ChromaDB ✅ Live
🔒 Hallucination Prevention Document answers grounded in retrieved chunks only ✅ Live
🤝 Multi-Agent System Specialized agents working in parallel 🔜 Phase 5
🧩 Multi-PDF Support Compare multiple documents in one session 🔜 Phase 4

📁 Project Structure

bizradar-ai/
│
├── 🤖 agent.py                # ReAct agent — Groq LLM + parallel tool execution loop
├── 🖥️  app.py                  # CLI entry point + PDF ingestion trigger
├── 🧠 context_manager.py      # Conversation memory — last 6 turns sliding window
├── 🛠️  tools.py               # Tool layer — Tavily search + Gemini analysis (self-summarizing) + RAG search
├── 📋 tools_description.py    # Tool schemas for LLM tool-calling (JSON format)
├── 📝 prompts.py              # System prompt + output format templates
├── 🗄️  rag.py                  # RAG pipeline — ingest, embed, store, query (Stage 4)
│
├── 📁 database/
│   └── chroma_db/             # Persistent ChromaDB vector store (local, gitignored)
│
├── 📦 requirements.txt        # Python dependencies
├── 🔒 .env                    # API keys (never committed)
├── 🚫 .gitignore              # Ignores .env, database, and local files
├── 📖 README.md               # This file
├── 🧭 ROADMAP.md              # Phase-by-phase build and learning path
├── 🏗️  ARCHITECTURE.md         # Deep dive into every design decision
├── 📓 LEARNING_LOG.md         # Personal learning tracker and mistake log
└── 📋 CHANGELOG.md            # Version history

🔁 ReAct Pattern — How The Agent Thinks

ReAct Pattern Architecture

🏗️ System Architecture

┌─────────────────────────────────────────────┐
│                   USER INPUT                │
│         "AI tiffin service for students"    │
└──────────────────────┬──────────────────────┘
                       │
                       ▼
┌─────────────────────────────────────────────┐
│           CLI INTERFACE  (app.py)           │
│     PDF ingestion trigger at startup        │
└──────────────────────┬──────────────────────┘
                       │
                       ▼
┌──────────────────────────────────────────────────────┐
│              ReAct AGENT LOOP  (agent.py)            │
│                                                      │
│  ┌──────────────┐     ┌────────────────────────────┐ │
│  │   Context    │     │       Groq LLM             │ │
│  │   Manager   │────▶│    Llama 3.3 70B           │ │
│  │  Last 6 msgs │     │  Reason → Act → Observe   │ │
│  └──────────────┘     └────────────┬───────────────┘ │
│                                    │                  │
│                    ┌───────────────▼─────────────┐   │
│                    │    ThreadPoolExecutor        │   │
│                    │    (Parallel Execution)      │   │
│                    │                             │   │
│                    │  analyze_market()            │   │
│                    │  search_knowledge_base()     │   │
│                    │  suggest_mvp()               │   │
│                    │  recommend_tech_stack()      │   │
│                    │  risk_analysis()             │   │
│                    │  search_documents() ← RAG    │   │
│                    │                             │   │
│                    │  Tavily · Gemini · ChromaDB  │   │
│                    └───────────────┬─────────────┘   │
└────────────────────────────────────┼─────────────────┘
                                     │
               ┌─────────────────────┴──────────────────┐
               │                                        │
               ▼                                        ▼
┌──────────────────────────┐           ┌───────────────────────────┐
│   STRUCTURED REPORT      │           │   RAG PIPELINE (rag.py)   │
│                          │           │                           │
│  # Market Potential      │           │  ingest_pdf()             │
│  # Competitors           │           │  embed_and_store()        │
│  # Suggested MVP         │           │  query_rag()              │
│  # Tech Stack            │           │                           │
│  # Risks                 │           │  ChromaDB Persistent      │
│  # Cited Sources         │           │  gemini-embedding-001     │
└──────────────────────────┘           └───────────────────────────┘

⚙️ Tech Stack

Layer Technology Purpose
Language Python 3.10+ Core runtime
LLM Inference Groq — Llama 3.3 70B Ultra-fast ReAct reasoning via LPU
Analysis Tools Gemini 2.5 Flash MVP, tech stack, risk generation
Embeddings Google gemini-embedding-001 Text → vectors for semantic search
Web Search Tavily API Real-time market research + citations
Vector Store ChromaDB Persistent Document chunk storage and retrieval
PDF Parsing pdfplumber Text extraction from pitch decks
Parallel Execution ThreadPoolExecutor Simultaneous tool execution
Memory Sliding window list Conversation context — last 6 turns
Tool Schemas JSON function definitions LLM tool-calling interface
Config python-dotenv Environment variable management

🚀 Getting Started

Prerequisites

1. Clone the Repository

git clone https://github.com/ankush-poonia007/STARTUP-AI-AGENT.git
cd STARTUP-AI-AGENT

2. Install Dependencies

pip install -r requirements.txt

3. Configure Environment Variables

# Create .env file in project root
GROQ_API_KEY=your_groq_key_here
TAVILY_API_KEY=your_tavily_key_here
GEMINI_API_KEY=your_gemini_key_here

⚠️ Never commit your .env file. It is gitignored by default.

4. Run BizRadar AI

python app.py

💬 Example Session

🚀 BizRadar AI Started
Type 'exit' to quit.

Do you have a document to upload? (YES / NO): yes
Enter Your File Path: ./pitch_deck.pdf
✅ Data ingestion complete. Data saved successfully.

You: What is the revenue projection in my pitch deck?

⚙️  Stage 1 of 3 — Executing tools...
   🔧 analyze_market()
   🔧 search_knowledge_base()
   ✅ Stage 1 complete.

⚙️  Stage 2 of 3 — Executing tools...
   🔧 suggest_mvp()
   🔧 recommend_tech_stack()
   ✅ Stage 2 complete.

⚙️  Stage 3 of 3 — Executing tools...
   🔧 risk_analysis()
   ✅ Stage 3 complete.

🔍 Stage 4 — Querying your document...
   🔧 search_documents()
   ✅ Report ready.


📊 BizRadar AI:

## Market Potential
High demand in tier-1 college cities. [Source: economictimes.com]

## Suggested MVP
- User registration & meal preferences
- Daily menu with AI personalization
- Subscription billing

## Recommended Tech Stack
- Backend: FastAPI  |  Frontend: React
- Database: PostgreSQL  |  AI Layer: Gemini API

## Risks
- High CAC in student market
- Competition from Swiggy/Zomato

## Final Verdict
Viable niche opportunity with strong MVP potential.

🧭 Roadmap

Phase Title Status
Phase 1 Foundation Agent — Ollama + Context + Tools ✅ Complete
Phase 2 Real Tools — Groq + ReAct + Parallel Execution ✅ Complete
Phase 3 RAG Pipeline — PDF Ingestion + ChromaDB + Vector Search ✅ Complete
Phase 4 Multi-PDF + Advanced RAG + Evaluation 🔜 Next
Phase 5 Multi-Agent Architecture 📋 Planned
Phase 6 Autonomous Research Platform 📋 Planned

🏛️ Design Philosophy

"Architecture First. Frameworks Later."

BizRadar is intentionally built without LangChain or LlamaIndex to deeply understand how AI agents work at the foundational level. Every component — the ReAct loop, tool calling, parallel execution, RAG pipeline, vector store — is implemented manually.

When you eventually use a framework, you will understand exactly what it is abstracting and why.


🎯 Learning Objectives

  • ✅ Prompt Engineering & System Design
  • ✅ Context Window Management
  • ✅ Tool-Augmented AI Agents
  • ✅ ReAct Agent Pattern
  • ✅ Parallel Tool Execution
  • ✅ OOP Architecture for AI Systems
  • ✅ Multi-Provider LLM Integration
  • ✅ Vector Embeddings & Semantic Search
  • ✅ RAG Pipeline from Scratch
  • ✅ ChromaDB & Persistent Vector Storage
  • 🔜 Multi-Document RAG & Evaluation
  • 📋 Multi-Agent Orchestration
  • 📋 Production AI Engineering

📜 License

MIT License — free to use, modify, and build on.


Built as an AI Engineering learning project — no frameworks, full understanding.

Star this repo if you found it useful.


👤 Contact

Ankush Poonia


B.Tech AI/ML · Arya College of Engineering · Jaipur


GitHub
LinkedIn
Email

Reviews (0)

No results found