MerchantryTidbits

Public catalog

Browse tools

27 shown. Search ranks against problem language; browsing defaults to stronger public signals.

rag-retrievallibraryfree

dspy

optimize LLM prompts and pipelines programmatically

Verified 2026-08-10 · 12,108 stars

evals-testingclifree

langfuse

observe LLM traces in production

Verified 2026-08-10 · 6,855 stars

evals-testingclifree

promptfoo

regression test LLM prompts and RAG outputs in CI

Verified 2026-08-10 · 5,452 stars

evals-testingpatternfree

prove-it-side-effect-check

require agents to prove side effects before claiming done

Verified 2026-08-10 · 5,452 stars

evals-testinglibraryfree

deepeval

pytest style LLM evaluation metrics

Verified 2026-08-10 · 3,114 stars

rag-retrievallibraryfree

TruLens

need offline evals for LLM app quality

Verified 2026-08-10 · 1,647 stars

evals-testinglibraryfree

OpenLLMetry

OpenTelemetry traces for LLM and agent calls

Verified 2026-08-10 · 1,587 stars

evals-testinglibraryfree

Weave

trace multi-step LLM apps with W&B Weave

Verified 2026-08-10 · 545 stars

frontend-toolinglibraryfree

axe-core

add automated accessibility assertions to browser and component tests

Verified 2026-08-10

observabilityclifree

betterstack-collector

collect Kubernetes or Docker logs metrics and traces without adding application instrumentation

Verified 2026-08-10

frontend-toolingclifree

Biome

use one fast formatter and linter across JavaScript TypeScript JSON and CSS

Verified 2026-08-10

api-toolingappfree

Bruno

test HTTP APIs with a Git-friendly desktop and CLI client

Verified 2026-08-10

evals-testinglibraryfree

Cypress

write interactive end-to-end and component tests for web applications

Verified 2026-08-10

mobile-toolinglibraryfree

Detox

run synchronized end-to-end tests for a React Native app

Verified 2026-08-10

evals-testingskillfree

Eval Skills: Eval Audit

My LLM evaluation pipeline has gaps and I need a prioritized audit of likely problems.

Verified 2026-08-10

observabilityclifree

k6

load test HTTP APIs with staged traffic and explicit thresholds

Verified 2026-08-10

mobile-toolingclifree

ktlint

enforce Kotlin style without maintaining a custom ruleset

Verified 2026-08-10

rag-retrievallibraryfree

OpenInference

standardize LLM span attributes across frameworks

Verified 2026-08-10

ml-operationslibraryfree

Optuna

run resumable hyperparameter optimization with pruning

Verified 2026-08-10

api-toolinglibraryfree

Pact JS

prevent breaking API changes with consumer-driven contracts

Verified 2026-08-10

evals-testinglibraryfree

phoenix

local tracing eval UI for LLM agent pipelines

Verified 2026-08-10

codebase-intelclilocal

pyrefly

type-check a large Python codebase with a fast CLI

Verified 2026-08-10

deploy-infraappfree

Renovate

automate dependency update pull requests across repositories

Verified 2026-08-10

frontend-toolingappfree

Storybook

develop and document UI components in isolated states

Verified 2026-08-10

mobile-toolingclifree

SwiftFormat

enforce deterministic Swift formatting across editors and CI

Verified 2026-08-10

frontend-toolinglibraryfree

Vitest

run fast unit and component tests in a Vite or TypeScript project

Verified 2026-08-10

mobile-toolingclifree

XcodeGen

generate a reproducible Xcode project from a reviewable specification

Verified 2026-08-10