★ 0 · Updated 2026-09-24
Provides structured generation and high-performance serving for LLMs and VLMs using RadixAttention prefix caching.
Browse skills that produce this output.
★ 0 · Updated 2026-09-24
Provides structured generation and high-performance serving for LLMs and VLMs using RadixAttention prefix caching.
★ 0 · Updated 2026-09-22
High-throughput LLM serving engine supporting OpenAI compatible API, quantization, and tensor parallelism.
★ 1 · Updated 2026-09-20
Official skill for integrating Firebase AI Logic Gemini API into applications, covering setup, multimodal inference, structured output, and security.
★ 0 · Updated 2026-09-20
Integrates Firebase AI Logic to call Gemini models from client SDKs for multimodal generation, streaming, and structured output.
★ 0 · Updated 2026-09-20
Integrates Firebase AI Logic and Gemini API into client apps for text generation, multimodal inference, streaming, and security.
★ 18 · Updated 2026-09-09
Invoke Google Gemini models for text generation, reasoning, and code tasks using the Python google-genai SDK.
★ 58 · Updated 2026-06-30
Guide usage of the Gemini API on Agent Platform with the Google Gen AI SDK across Python, JS/TS, Go, Java, and C#.
★ 0 · Updated 2026-05-11
Access Cohere models for embeddings, reranking, RAG, and text generation via API