Fireworks AI

Serverless LLM inference with fine-tuning, RAG support, and free credits for rapid prototyping