Primitive Explorer · 847 primitives
Primitives
Every authoring primitive across the FrootAI catalog. Agents, skills, instructions, hooks, plugins, workflows — searchable + faceted. Each one is MIT-licensed and runnable today.
Filter
26 of 847 primitivesClear filters
- 🤖 agentsused by 2
fai-ai-infra-expert
AI infrastructure expert — GPU compute sizing (A100/H100), VRAM estimation, model serving (vLLM/TensorRT-LLM/Triton), AKS node pool design, PTU vs PAYG cost modeling, and quantization strategies.
agents/fai-ai-infra-expert.agent.md107 lines- performance-efficiency
- cost-optimization
- reliability
… - 🤖 agentsused by 2
fai-api-gateway-designer
API gateway architect — Azure APIM patterns, rate limiting, token-based throttling, multi-region load balancing, backend circuit breakers, and cost-aware routing for LLM endpoints.
agents/fai-api-gateway-designer.agent.md80 lines- cost-optimization
- reliability
- performance-efficiency
… - 🤖 agentsused by 2
fai-architect
Senior cloud-native solution architect — Azure Well-Architected Framework alignment, AI system design, multi-service integration, cost modeling, trade-off analysis, and production readiness assessment.
agents/fai-architect.agent.md78 lines- reliability
- security
- cost-optimization
- performance-efficiency
- +2
… - 🤖 agentsused by 6
fai-azure-ai-search-expert
Azure AI Search specialist — HNSW vector indexes, hybrid keyword+vector retrieval, semantic ranker, integrated vectorization pipelines, custom skillsets, scoring profiles, and RAG optimization for production search experiences.
agents/fai-azure-ai-search-expert.agent.md122 lines- performance-efficiency
- reliability
- cost-optimization
- security
… - 🤖 agentsused by 3
fai-azure-aks-expert
Azure Kubernetes Service specialist — GPU node pools (A100/H100), NVIDIA device plugin, model serving with vLLM/TGI/Triton, HPA/KEDA autoscaling, and production AI inference workload patterns.
agents/fai-azure-aks-expert.agent.md139 lines- performance-efficiency
- reliability
- cost-optimization
- security
… - 🤖 agentsused by 2
fai-azure-apim-expert
Azure API Management specialist — AI Gateway patterns, semantic caching, token metering, multi-backend load balancing, circuit breaker, rate limiting, and FinOps for LLM API layers.
agents/fai-azure-apim-expert.agent.md143 lines- cost-optimization
- reliability
- performance-efficiency
- security
… - 🤖 agentsused by 4
fai-azure-cosmos-db-expert
Azure Cosmos DB specialist — partition key design, DiskANN vector search, multi-region writes, RU optimization, change feed processing, and conversation/session storage for AI agents.
agents/fai-azure-cosmos-db-expert.agent.md146 lines- performance-efficiency
- reliability
- cost-optimization
- security
… - 🤖 agentsused by 3
fai-azure-event-hubs-expert
Azure Event Hubs specialist — partitioned event streaming, Kafka compatibility, Schema Registry governance, real-time AI inference pipelines, and high-throughput data ingestion.
agents/fai-azure-event-hubs-expert.agent.md163 lines- performance-efficiency
- reliability
- cost-optimization
… - 🤖 agentsused by 3
fai-azure-openai-expert
Azure OpenAI specialist — model deployment types (PTU/PAYG/Global), content filtering, structured output, token optimization, multi-region load balancing, and production inference patterns.
agents/fai-azure-openai-expert.agent.md164 lines- cost-optimization
- performance-efficiency
- security
- responsible-ai
… - 🤖 agentsused by 2
fai-azure-sql-expert
Azure SQL specialist — Hyperscale, serverless auto-pause, native vector search, geo-replication, intelligent performance tuning, and AI integration patterns with embeddings storage.
agents/fai-azure-sql-expert.agent.md154 lines- performance-efficiency
- reliability
- cost-optimization
- security
… - 🤖 agentsused by 3
fai-batch-processing-expert
Batch processing specialist — Azure Batch pools, Global Batch API (50% cost savings), Durable Functions fan-out, large-scale document/embedding pipelines, and async LLM inference patterns.
agents/fai-batch-processing-expert.agent.md159 lines- cost-optimization
- reliability
- performance-efficiency
… - 🤖 agentsused by 3
fai-capacity-planner
AI capacity planning specialist — GPU sizing, PTU allocation, token volume forecasting, cost modeling, scaling strategy, and FinOps for Azure AI workloads.
agents/fai-capacity-planner.agent.md119 lines- cost-optimization
- performance-efficiency
- reliability
… - 🤖 agentsused by 2
fai-cost-optimizer
FinOps cost optimizer for AI workloads — model routing economics, semantic caching ROI, token budget design, PTU vs PAYG analysis, right-sizing recommendations, and Azure cost attribution.
agents/fai-cost-optimizer.agent.md147 lines- cost-optimization
- performance-efficiency
… - 🤖 agentsused by 2
fai-dspy-expert
DSPy framework specialist — declarative LM programs, signature-based modules, optimizers (BootstrapFewShot, MIPRO), assertions, metric-driven prompt optimization, and compiled prompt pipelines.
agents/fai-dspy-expert.agent.md155 lines- performance-efficiency
- cost-optimization
- reliability
… - 🤖 agentsused by 2
fai-embedding-expert
Embedding specialist — text-embedding-3 model selection, Matryoshka dimension reduction, batch embedding pipelines, similarity metrics, chunking strategies, and vector database integration for RAG.
agents/fai-embedding-expert.agent.md150 lines- cost-optimization
- performance-efficiency
… - 🤖 agentsused by 2
fai-event-driven-expert
Event-driven architecture specialist — Azure Event Grid, Service Bus, Event Hubs selection, event sourcing, CQRS, saga orchestration, and exactly-once processing patterns for AI pipelines.
agents/fai-event-driven-expert.agent.md143 lines- reliability
- performance-efficiency
- cost-optimization
… - 🤖 agentsused by 2
fai-genai-foundations-expert
GenAI foundations expert — transformer architecture, tokenization, inference optimization (KV cache, speculative decoding), model taxonomy, prompt engineering theory, and evaluation benchmarks.
agents/fai-genai-foundations-expert.agent.md123 lines- performance-efficiency
- cost-optimization
… - 🤖 agentsused by 2
fai-llm-landscape-expert
LLM landscape expert — model families (GPT, Claude, Llama, Gemini, Phi), benchmarks (MMLU, HumanEval, MT-Bench), deployment types, quantization, and model selection frameworks.
agents/fai-llm-landscape-expert.agent.md118 lines- cost-optimization
- performance-efficiency
… - 🤖 agentsused by 2
fai-ml-engineer
ML engineering specialist — model training pipelines, LoRA/QLoRA fine-tuning, evaluation metrics, MLOps with Azure AI Foundry, model registry, and serving optimization for production AI.
agents/fai-ml-engineer.agent.md144 lines- cost-optimization
- performance-efficiency
- operational-excellence
… - 🤖 agentsused by 2
fai-performance-profiler
Performance profiling specialist — latency analysis (P50/P95/P99), token optimization, GPU utilization profiling, bottleneck identification, cold start analysis, and AI pipeline performance tuning.
agents/fai-performance-profiler.agent.md149 lines- performance-efficiency
- cost-optimization
… - 🤖 agentsused by 2
fai-production-patterns-expert
Production AI patterns expert — hosting selection (Container Apps/AKS/Functions), APIM gateway patterns, streaming SSE, retry/circuit-breaker, health checks, and multi-region deployment for production AI workloads.
agents/fai-production-patterns-expert.agent.md175 lines- reliability
- performance-efficiency
- cost-optimization
- operational-excellence
… - 🤖 agentsused by 2
fai-rag-architect
Enterprise RAG architecture specialist — designs end-to-end retrieval-augmented generation pipelines with Azure AI Search, OpenAI embeddings, chunking strategies, grounding, citation, evaluation, and production deployment.
agents/fai-rag-architect.agent.md144 lines- security
- reliability
- cost-optimization
- performance-efficiency
… - 🤖 agentsused by 3
fai-rag-expert
RAG expert — advanced retrieval patterns (agentic, graph, multi-modal RAG), chunking strategies, hybrid search, re-ranking, evaluation metrics, and production RAG optimization.
agents/fai-rag-expert.agent.md169 lines- reliability
- performance-efficiency
- cost-optimization
… - 🤖 agentsused by 2
fai-ray-expert
Ray distributed computing specialist — Ray Serve for model serving, Ray Tune for hyperparameter optimization, Ray Data for preprocessing, and distributed training/inference at scale.
agents/fai-ray-expert.agent.md148 lines- performance-efficiency
- cost-optimization
…