-
strands-agents-langfuse-evaluations
🏦 PoC using Strands Agents with Langfuse tracing, offline/online evaluations, prompt management, and annotation queues.
See source code at Github
-
SaaSHub
SaaSHub - Software Alternatives and Reviews. SaaSHub helps you find the best software and product alternatives
-
harness-sdk
Build an agent harness and control it end-to-end. Open-source SDK for production AI agents in Python & TypeScript - any model, any cloud.
Strands Agents is a lightweight Python SDK for building LLM-powered agents with tool use and session memory, open-sourced by AWS in May 2025. It is Python-native — which pairs well with the Langfuse Python SDK — and new enough to be worth exploring. Any other Python agent framework would work just as well for this PoC.
-
langfuse
🪢 Open source AI engineering platform: LLM evals, observability, metrics, prompt management, playground, datasets. Integrates with OpenTelemetry, LangChain, OpenAI SDK, LiteLLM, and more. 🍊YC W23
In this project we will build a Python banking assistant agent using Strands Agents and make it observable and continuously evaluated using Langfuse — step by step.
-
That's where traces and evaluations come in — supported by a growing number of platforms: Langfuse, Arize Phoenix, MLflow, LangSmith, W&B Weave, Datadog, AWS AgentCore, Azure AI Foundry, Google Vertex AI
Related posts
-
Long-Text Summarization Explained: Cheap Node.js API Token Counts for SaaS
-
How to Benchmark a Unified LLM API: Node.js Invoice Extraction Under Latency Limits
-
Tenant Cost Attribution Explained — 3 Fallback Models Behind One Chatbot API
-
Per-Tenant Media Scheduling Explained: Node.js LLM JSON Schema for Ticket Tags
-
Speech-to-Text API 429 Triage: Backoff, Queue, and Batch Transcription