Groq

[WHAT] Fast LLM inference via Groq API (chat) + Ollama (embeddings) [HOW] Groq for chat completions (llama-3.3-70b-versatile), Ollama nomic-embed-text for embeddings [WHEN] Need fast inference, embedding text for RAG, chat completions [WHY] Groq provides fastest LLM inference; Ollama handles local embeddings (Groq has no embedding API) Triggers: "groq embed", "groq chat", "groq complete", "embed with groq", "fast llm"

lev-os 23cf45e 4 files · 7.2 KB Updated

File contents

lev-os/agents/tree/main/skills-db/sdk/groq commit 23cf45e07d

Frequently asked questions

npx skillmds@latest add lev-os/groq