Skip to content
View roydonsequeira's full-sized avatar
💭
Working from home
💭
Working from home

Highlights

  • Pro

Block or report roydonsequeira

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
roydonsequeira/README.md
Roydon Sequeira, GenAI Engineer and AI Agent Architect. I build AI agents that plan, retrieve and act, and know when to hand off to a human.

Portfolio: roydonsequeira.com LinkedIn X: @roydonsequeiraa Email

whoami  open source  recently shipped  how I build agents  experience  stack  activity

whoami

Role: AI Agent Developer at Code Crew Studio. Previously GenAI Engineer at ReinHealth, a stealth AI healthcare startup. Focus: multi-agent systems, LLM orchestration, RAG. Stack: Python, LangChain, Qdrant, FastAPI, Docker. Runtime: Ollama for local-first, private inference. 2+ years shipping production AI. Based in Udupi, Karnataka, India. Open to AI engineering roles and contracts. Impact: about 70% less manual patient-intake workload; patient intake down from 15 minutes to under 5; zero patient records sent to external APIs; 391 tests behind CORTEX (204 unit plus 187 live conversations).

Intake figures come from the production medical agent I led at ReinHealth. Test counts come from CORTEX's unit suite and live test battery.

Open source

Local-first by default. Most of this runs entirely on hardware you own, with no API keys.

CORTEX, Private Intelligence Framework: a private AI agent that runs entirely on your own machine, with planning, sandboxed tools, four-tier memory, LATS tree search, a streaming UI and OpenTelemetry tracing, powered by Ollama.

CI status Last commit Tests: 204 unit and 187 live License: MIT

How CORTEX works: the architecture, the guardrails, and a quick start
flowchart LR
    UI["Next.js UI"] -->|SSE| API["FastAPI gateway"]
    API --> K["Agent kernel<br/>plan · act · reflect"]
    K -. hard tasks .-> LATS["LATS tree search"]
    K -. big jobs .-> SUP["Supervisor → workers"]
    K --> MEM["Memory<br/>working · episodic · semantic · procedural"]
    K --> TOOLS["Sandboxed tools<br/>python · files · web · docs"]
    K --> R["Model router"] --> O["Ollama"]
    API -. OTLP spans .-> J["Jaeger"]
Loading

Why it behaves on a 7B model. Small models break rules that prompts ask them to follow, so CORTEX enforces the rules in code:

  • Plans that need no tools run with no tool schemas at all, and destructive plans become refusals before anything executes.
  • Pasted or quoted text is treated as data, so instructions hidden inside it can't trigger a tool.
  • Web fetch reaches public hosts only. Loopback, private-network and cloud-metadata addresses are refused after DNS resolution and on every redirect.
  • Memory stores only durable facts you state about yourself.
git clone https://github.com/roydonsequeira/CORTEX-Private-Intelligence-Framework.git && cd CORTEX-Private-Intelligence-Framework
ollama pull qwen2.5:7b && ollama pull nomic-embed-text
pip install -e . && cortex doctor && cortex serve    # API on :8000
cd ui && npm install && npm run dev                   # UI on :3000

RagChatbot: chat with your own PDFs, scans and Word files entirely on your machine, with OCR ingestion, local embeddings and streamed answers. Clinical SOAP Notes: n8n workflows that turn clinical text or voice into structured SOAP notes with a local LLM. Skin Lesion Segmentation: U-Net in PyTorch and ResU-Net in TensorFlow 2 on 2,594 ISIC 2018 dermoscopy images, with a Streamlit demo. In the lab, private builds: NEXUS AI content intelligence SaaS; a multi-agent code review agent; an AI resume analyzer with 3 agents, under 400 ms and 39 tests; an AI proxy agent for WhatsApp, calendar and leads; a research agent with search, retrieval and citations.

Recently shipped

Project What it is Last push
CORTEX-Private-Intelligence-Framework
Python · ★ 4
Private, local-first AI agent: planning, sandboxed tools, four-tier memory, streaming UI and OpenTelemetry — runs entirely on your machine with Ollama. 2026-10-07
Skin-Lesion-Segmentation-in-TensorFlow-2.0
Python
Skin lesion segmentation on ISIC 2018 using U-Net (PyTorch) and ResU-Net (TensorFlow 2). Final year project. 2026-05-25
RagChatbot
Python
Production-quality RAG chatbot: Ollama LLM, ChromaDB vector store, OCR/PDF ingestion, Next.js UI 2026-04-09
clinicalNote-SOAP n8n workflows for AI-powered Clinical SOAP note generation (TTT, STT, TTS) 2026-03-25

Refreshed daily from the GitHub API by readme-sync. New public repositories show up here on their own.

How I build agents

flowchart LR
    IN["User or trigger"] --> ORCH["Orchestrator"]
    ORCH --> PLAN["Plan and reason"]
    ORCH --> RET["Retrieve context"]
    ORCH --> TOOLS["Call tools"]
    RET --> VEC["Qdrant / ChromaDB"]
    TOOLS --> SYS["APIs / databases"]
    PLAN --> CHECK{"Validate and guardrail"}
    VEC --> CHECK
    SYS --> CHECK
    CHECK -->|pass| OUT["Answer or action"]
    CHECK -->|fail| ESC["Fallback or human handoff"]
Loading
Principle What it looks like in production
Rules live in code, not prompts CORTEX turns destructive plans into refusals before any tool runs, and pasted text can never trigger an action.
Keep the data home Local inference on Ollama: the ReinHealth intake agent sent zero patient records to external APIs.
Ground every answer Retrieval over Qdrant, ChromaDB and PostgreSQL instead of trusting model memory.
Escalate instead of guessing Emergency-symptom escalation, input validation and audit logging in the clinical agent.
Trace everything OpenTelemetry spans across model, tool and memory calls, so a failure is debugged as a system rather than guessed at as a prompt.

Experience

Career timeline: ML Intern at Igeeks Technologies (Jun to Jul 2023); GenAI Engineer at ReinHealth (Jul 2024 to Feb 2026); AI Agent Developer at Code Crew Studio (Feb 2026 to now). B.E. in AI and Machine Learning at NMAM Institute of Technology (2020 to 2024); Executive PG Certification in Data Science and AI, iHUB DivyaSampark, IIT Roorkee (2024 to 2026).
AI Agent Developer · Code Crew Studio · Feb 2026 – present · Mumbai (remote)
  • Built and maintain a production onboarding agent on the official WhatsApp Business API that runs structured intake conversations, classifies user problems and returns analyzed feedback.
  • Engineered a multi-turn query system with persistent context across sessions, using LangChain for orchestration-heavy flows and direct LLM API calls where latency matters.
  • Built FastAPI services on PostgreSQL for agent state, conversation history and structured customer records.
GenAI Engineer · ReinHealth, stealth AI healthcare startup · Jul 2024 – Feb 2026 · Colorado, US (remote)
  • Led end-to-end delivery of a production autonomous medical intake agent over text and voice, cutting manual intake work by an estimated 70% and average intake time from 15 minutes to under 5.
  • Architected an LLM orchestration layer on local Ollama inference (intent analysis, follow-up selection, tool execution, grounded synthesis), so zero patient records reached external APIs.
  • Built a privacy-first RAG pipeline on Qdrant with PostgreSQL, and automated speech-to-text clinical notes, TTS summaries and conflict-aware scheduling in n8n.
  • Hardened the platform for clinical use with input validation, emergency-symptom escalation and audit logging, behind a single REST API for agents, workflows and frontend.
Machine Learning Intern · Igeeks Technologies · Jun – Jul 2023 · Bengaluru
  • Built CNN, AlexNet and MLP image-classification pipelines on custom datasets, with OpenCV preprocessing and hyperparameter tuning.

Education: B.E. in Artificial Intelligence & Machine Learning, NMAM Institute of Technology (2020–2024) · Executive PG Certification in Data Science & AI, iHUB DivyaSampark, IIT Roorkee (2024–2026)

Stack

Layer Tools
Agents & LLMs LangChain · ReAct · LATS · supervisor–worker orchestration · Ollama · Claude · Gemini · OpenAI · Hugging Face · STT/TTS
Retrieval Qdrant · ChromaDB · BGE and nomic embeddings · semantic search · PostgreSQL · Redis
Backend Python · FastAPI · Flask · Pydantic · SQL · REST · SSE streaming
Automation n8n · WhatsApp Business API · Twilio
ML & vision PyTorch · TensorFlow · scikit-learn · OpenCV · NumPy · Pandas
Ship & observe Docker · Linux · GitHub Actions · OpenTelemetry · Vercel · ruff · mypy
Agent UIs Next.js · TypeScript · Tailwind CSS

Python, FastAPI, Flask, PostgreSQL, Redis, Docker, Linux, GitHub Actions, PyTorch, TensorFlow, scikit-learn, OpenCV, Next.js, TypeScript, Tailwind CSS, Vercel

LangChain Ollama Qdrant n8n Claude Gemini Hugging Face OpenTelemetry WhatsApp Business API

Activity

Contribution graph being eaten by a snake

GitHub contribution streak


Let's build agents that ship. Open to AI engineering roles, contracts and production agent builds.

Email me Portfolio

Profile views

Popular repositories Loading

  1. CORTEX-Private-Intelligence-Framework CORTEX-Private-Intelligence-Framework Public

    Private, local-first AI agent: planning, sandboxed tools, four-tier memory, streaming UI and OpenTelemetry — runs entirely on your machine with Ollama.

    Python 4 1

  2. Sequeira-roy Sequeira-roy Public

    Config files for my GitHub profile.

  3. roydonsequeira roydonsequeira Public

    Python

  4. Skin-Lesion-Segmentation-in-TensorFlow-2.0 Skin-Lesion-Segmentation-in-TensorFlow-2.0 Public

    Skin lesion segmentation on ISIC 2018 using U-Net (PyTorch) and ResU-Net (TensorFlow 2). Final year project.

    Python

  5. clinicalNote-SOAP clinicalNote-SOAP Public

    n8n workflows for AI-powered Clinical SOAP note generation (TTT, STT, TTS)

  6. RagChatbot RagChatbot Public

    Production-quality RAG chatbot: Ollama LLM, ChromaDB vector store, OCR/PDF ingestion, Next.js UI

    Python