MT

Muhammad Tariq

    AI agent
    development.

    An AI agent is software that uses a large language model to decide which steps and tools to use to complete a task — for example looking up data, calling an API and drafting a reply. I build agents with tool calling, retrieval (RAG), memory, logging and human approval for high-risk actions.

    I build AI agents that complete real tasks inside your business — looking things up, calling your APIs, drafting responses and updating records — with guardrails, logs and human approval where the stakes are high.

    01

    What I build

    Tool-calling agents

    Agents that call your APIs, databases and third-party services with typed inputs.

    Multi-agent workflows

    Specialised agents for research, drafting and review, coordinated by a planner.

    RAG and memory

    Agents grounded in your documents, with short- and long-term memory.

    Human approval

    Approval queues before an agent sends, pays or deletes anything.

    02

    Problems I solve

    • Agents that loop, hallucinate tools or take unsafe actions
    • No record of why an agent did something
    • Vendor lock-in to one model provider
    03

    Technologies I use

    • OpenAI
    • Claude
    • Gemini
    • Grok
    • DeepSeek
    • Ollama
    • Structured outputs
    • Python
    • FastAPI
    • PostgreSQL + pgvector
    • Redis
    04

    Development process

    1. 01

      Discovery

    2. 02

      Architecture

    3. 03

      Build

    4. 04

      Test

    5. 05

      Deploy

    6. 06

      Iterate

    05

    Relevant projects

    06

    Technical approach

    • →OpenAI and Claude handle complex reasoning and tool calling where accuracy matters most.
    • →Gemini is used for long documents and multimodal inputs such as images and PDFs.
    • →Grok and DeepSeek are options for cost-sensitive or high-volume steps, chosen by measured quality.
    • →Ollama runs open models locally or on your servers when data must not leave your infrastructure.
    • →An LLM router picks the model per step, so switching providers is a config change, not a rewrite.
    • →Every tool call is logged with inputs and outputs, and risky actions pause for human approval.
    07

    Frequently asked questions

    Can an agent act on its own?

    For low-risk actions, yes. For anything involving money, customers or deletion, I add an approval step.

    Which models do you support?

    OpenAI, Claude, Gemini, Grok, DeepSeek and local models via Ollama, behind one routing layer.

    How long does a typical project take?

    It depends on scope. A focused feature or fix can take days; a first production version of a product usually takes several weeks. I give a written estimate after a short discovery call.

    08

    Related services

    09 — Next step

    Ready to start
    your project?

    I work remotely with startups and businesses across the United States and internationally. Tell me what you're building and I'll reply with next steps.