Skip to main content

Primitive Explorer · 847 primitives

Primitives

Every authoring primitive across the FrootAI catalog. Agents, skills, instructions, hooks, plugins, workflows — searchable + faceted. Each one is MIT-licensed and runnable today.

← Home

26 of 847 primitivesClear filters

  • 🤖 agentsused by 2

    fai-ai-infra-expert

    AI infrastructure expert — GPU compute sizing (A100/H100), VRAM estimation, model serving (vLLM/TensorRT-LLM/Triton), AKS node pool design, PTU vs PAYG cost modeling, and quantization strategies.

    agents/fai-ai-infra-expert.agent.md107 lines
    • performance-efficiency
    • cost-optimization
    • reliability
  • 🤖 agentsused by 2

    fai-api-gateway-designer

    API gateway architect — Azure APIM patterns, rate limiting, token-based throttling, multi-region load balancing, backend circuit breakers, and cost-aware routing for LLM endpoints.

    agents/fai-api-gateway-designer.agent.md80 lines
    • cost-optimization
    • reliability
    • performance-efficiency
  • 🤖 agentsused by 2

    fai-architect

    Senior cloud-native solution architect — Azure Well-Architected Framework alignment, AI system design, multi-service integration, cost modeling, trade-off analysis, and production readiness assessment.

    agents/fai-architect.agent.md78 lines
    • reliability
    • security
    • cost-optimization
    • performance-efficiency
    • +2
  • 🤖 agentsused by 6

    fai-azure-ai-search-expert

    Azure AI Search specialist — HNSW vector indexes, hybrid keyword+vector retrieval, semantic ranker, integrated vectorization pipelines, custom skillsets, scoring profiles, and RAG optimization for production search experiences.

    agents/fai-azure-ai-search-expert.agent.md122 lines
    • performance-efficiency
    • reliability
    • cost-optimization
    • security
  • 🤖 agentsused by 3

    fai-azure-aks-expert

    Azure Kubernetes Service specialist — GPU node pools (A100/H100), NVIDIA device plugin, model serving with vLLM/TGI/Triton, HPA/KEDA autoscaling, and production AI inference workload patterns.

    agents/fai-azure-aks-expert.agent.md139 lines
    • performance-efficiency
    • reliability
    • cost-optimization
    • security
  • 🤖 agentsused by 2

    fai-azure-apim-expert

    Azure API Management specialist — AI Gateway patterns, semantic caching, token metering, multi-backend load balancing, circuit breaker, rate limiting, and FinOps for LLM API layers.

    agents/fai-azure-apim-expert.agent.md143 lines
    • cost-optimization
    • reliability
    • performance-efficiency
    • security
  • 🤖 agentsused by 4

    fai-azure-cosmos-db-expert

    Azure Cosmos DB specialist — partition key design, DiskANN vector search, multi-region writes, RU optimization, change feed processing, and conversation/session storage for AI agents.

    agents/fai-azure-cosmos-db-expert.agent.md146 lines
    • performance-efficiency
    • reliability
    • cost-optimization
    • security
  • 🤖 agentsused by 3

    fai-azure-event-hubs-expert

    Azure Event Hubs specialist — partitioned event streaming, Kafka compatibility, Schema Registry governance, real-time AI inference pipelines, and high-throughput data ingestion.

    agents/fai-azure-event-hubs-expert.agent.md163 lines
    • performance-efficiency
    • reliability
    • cost-optimization
  • 🤖 agentsused by 3

    fai-azure-openai-expert

    Azure OpenAI specialist — model deployment types (PTU/PAYG/Global), content filtering, structured output, token optimization, multi-region load balancing, and production inference patterns.

    agents/fai-azure-openai-expert.agent.md164 lines
    • cost-optimization
    • performance-efficiency
    • security
    • responsible-ai
  • 🤖 agentsused by 2

    fai-azure-sql-expert

    Azure SQL specialist — Hyperscale, serverless auto-pause, native vector search, geo-replication, intelligent performance tuning, and AI integration patterns with embeddings storage.

    agents/fai-azure-sql-expert.agent.md154 lines
    • performance-efficiency
    • reliability
    • cost-optimization
    • security
  • 🤖 agentsused by 3

    fai-batch-processing-expert

    Batch processing specialist — Azure Batch pools, Global Batch API (50% cost savings), Durable Functions fan-out, large-scale document/embedding pipelines, and async LLM inference patterns.

    agents/fai-batch-processing-expert.agent.md159 lines
    • cost-optimization
    • reliability
    • performance-efficiency
  • 🤖 agentsused by 3

    fai-capacity-planner

    AI capacity planning specialist — GPU sizing, PTU allocation, token volume forecasting, cost modeling, scaling strategy, and FinOps for Azure AI workloads.

    agents/fai-capacity-planner.agent.md119 lines
    • cost-optimization
    • performance-efficiency
    • reliability
  • 🤖 agentsused by 2

    fai-cost-optimizer

    FinOps cost optimizer for AI workloads — model routing economics, semantic caching ROI, token budget design, PTU vs PAYG analysis, right-sizing recommendations, and Azure cost attribution.

    agents/fai-cost-optimizer.agent.md147 lines
    • cost-optimization
    • performance-efficiency
  • 🤖 agentsused by 2

    fai-dspy-expert

    DSPy framework specialist — declarative LM programs, signature-based modules, optimizers (BootstrapFewShot, MIPRO), assertions, metric-driven prompt optimization, and compiled prompt pipelines.

    agents/fai-dspy-expert.agent.md155 lines
    • performance-efficiency
    • cost-optimization
    • reliability
  • 🤖 agentsused by 2

    fai-embedding-expert

    Embedding specialist — text-embedding-3 model selection, Matryoshka dimension reduction, batch embedding pipelines, similarity metrics, chunking strategies, and vector database integration for RAG.

    agents/fai-embedding-expert.agent.md150 lines
    • cost-optimization
    • performance-efficiency
  • 🤖 agentsused by 2

    fai-event-driven-expert

    Event-driven architecture specialist — Azure Event Grid, Service Bus, Event Hubs selection, event sourcing, CQRS, saga orchestration, and exactly-once processing patterns for AI pipelines.

    agents/fai-event-driven-expert.agent.md143 lines
    • reliability
    • performance-efficiency
    • cost-optimization
  • 🤖 agentsused by 2

    fai-genai-foundations-expert

    GenAI foundations expert — transformer architecture, tokenization, inference optimization (KV cache, speculative decoding), model taxonomy, prompt engineering theory, and evaluation benchmarks.

    agents/fai-genai-foundations-expert.agent.md123 lines
    • performance-efficiency
    • cost-optimization
  • 🤖 agentsused by 2

    fai-llm-landscape-expert

    LLM landscape expert — model families (GPT, Claude, Llama, Gemini, Phi), benchmarks (MMLU, HumanEval, MT-Bench), deployment types, quantization, and model selection frameworks.

    agents/fai-llm-landscape-expert.agent.md118 lines
    • cost-optimization
    • performance-efficiency
  • 🤖 agentsused by 2

    fai-ml-engineer

    ML engineering specialist — model training pipelines, LoRA/QLoRA fine-tuning, evaluation metrics, MLOps with Azure AI Foundry, model registry, and serving optimization for production AI.

    agents/fai-ml-engineer.agent.md144 lines
    • cost-optimization
    • performance-efficiency
    • operational-excellence
  • 🤖 agentsused by 2

    fai-performance-profiler

    Performance profiling specialist — latency analysis (P50/P95/P99), token optimization, GPU utilization profiling, bottleneck identification, cold start analysis, and AI pipeline performance tuning.

    agents/fai-performance-profiler.agent.md149 lines
    • performance-efficiency
    • cost-optimization
  • 🤖 agentsused by 2

    fai-production-patterns-expert

    Production AI patterns expert — hosting selection (Container Apps/AKS/Functions), APIM gateway patterns, streaming SSE, retry/circuit-breaker, health checks, and multi-region deployment for production AI workloads.

    agents/fai-production-patterns-expert.agent.md175 lines
    • reliability
    • performance-efficiency
    • cost-optimization
    • operational-excellence
  • 🤖 agentsused by 2

    fai-rag-architect

    Enterprise RAG architecture specialist — designs end-to-end retrieval-augmented generation pipelines with Azure AI Search, OpenAI embeddings, chunking strategies, grounding, citation, evaluation, and production deployment.

    agents/fai-rag-architect.agent.md144 lines
    • security
    • reliability
    • cost-optimization
    • performance-efficiency
  • 🤖 agentsused by 3

    fai-rag-expert

    RAG expert — advanced retrieval patterns (agentic, graph, multi-modal RAG), chunking strategies, hybrid search, re-ranking, evaluation metrics, and production RAG optimization.

    agents/fai-rag-expert.agent.md169 lines
    • reliability
    • performance-efficiency
    • cost-optimization
  • 🤖 agentsused by 2

    fai-ray-expert

    Ray distributed computing specialist — Ray Serve for model serving, Ray Tune for hyperparameter optimization, Ray Data for preprocessing, and distributed training/inference at scale.

    agents/fai-ray-expert.agent.md148 lines
    • performance-efficiency
    • cost-optimization

Catalog synthesised from frootai/. Refresh with pnpm catalog:sync. Default page size: 24.