pydantic/pydantic-ai
pydantic/pydantic-ai is an MIT-licensed Python agent framework from the Pydantic team that gives developers a type-safe, model-agnostic, FastAPI-style way to build production LLM…
RepoRadar helps AI developers, technical founders, and engineering leads answer three questions: What changed? What should I try? Which option fits this workflow?
Saved tools, searches, risks, movers, and newly reviewed candidates appear here first.
Check the newest retained change window when you are ready. Nothing is loaded until you ask.
Open the full change history →Today's rotating evidence-backed picks, selected from retained editorial or hands-on review. Each card shows its actual verification stage and opens the RepoRadar record first.
pydantic/pydantic-ai is an MIT-licensed Python agent framework from the Pydantic team that gives developers a type-safe, model-agnostic, FastAPI-style way to build production LLM…
OpenAI Agents SDK for JavaScript/TypeScript is an MIT-licensed, provider-agnostic framework for building multi-agent and voice-agent workflows with tools, MCP support, handoffs…
microsoft/presidio is an MIT-licensed, context-aware PII de-identification SDK from Microsoft that detects and anonymizes sensitive entities in text and images and ships pluggable…
Fission-AI/OpenSpec is an MIT spec-driven development framework and CLI for AI coding assistants that turns vague requests into reviewable proposals, plans, task artifacts, and…
neuml/txtai is an Apache-2.0-licensed, open-source all-in-one AI framework for semantic search, LLM orchestration, embeddings, and language-model workflows that ships as a Python…
MIT Microsoft cross-platform ML inference and training accelerator with ONNX standard support + hardware accelerators across CPU / GPU / NPU; 21,042* at verify time, last commit…
Repeat-visitor view: new signals, score movement, risk-label changes, and fresh try-now picks from the last 24 hours.
Tools, repos, agents, apps, and frameworks added in the last 24 hours.
No change detected in the last 24 hours.
No change detected in the last 24 hours.
No change detected in the last 24 hours.
No change detected in the last 24 hours.
Newly added or newly promoted try-now verdicts in the last 24 hours.
llama.cpp is an MIT-licensed C/C++ runtime for local LLM inference with GGUF models, server builds, Docker packaging, and a very large maintainer/user footprint. Current GitHub met
Genesis World is an Apache-2.0 simulation platform for robotics and embodied AI that combines a Python simulation interface, multi-physics engine, robotics-focused rendering, and a
Garak is NVIDIA's Apache-2.0 LLM red-team scanner. The current README describes probes for hallucination, data leakage, prompt injection, misinformation, toxicity, jailbreaks, and
Axolotl is an Apache-2.0 framework for fine-tuning language models. The current README describes it as a free and open-source LLM fine-tuning framework, while GitHub metadata and t
LangExtract is an Apache-2.0 Python library for turning long unstructured text into structured data with exact source grounding, schema control, and an interactive viewer for inspe
Reviewed decision records from this run. Each passed the publication gate and links to its evidence and verification basis. Rotates weekly so returning visitors see fresh picks.
Why it's featured: it cleared our evidence and quality gates with the strongest overall signal in this run — the pick we’d point a newcomer to first.
Open decision record →Tested in a bounded workflowGitHub source ↗ for langchain-ai/langchainMomentum is a research lead, not a recommendation or editorial verification. Trending reflects engagement velocity and recency; each card shows its actual verification stage when one exists.
earendil-works/pi is an MIT agent toolkit that bundles a unified multi-provider LLM API, agent runtime, interactive coding-agent CLI, terminal UI…
Tested in a bounded workflow GitHub source ↗ for earendil-works/piopen-code-review alibaba/open-code-review is an Apache-2.0 open-source AI-powered code-review CLI that originated as Alibaba Group's internal…
Research signal · Discovered, not yet verified GitHub source ↗ for alibaba/open-code-reviewCaveman JuliusBrussee/caveman is an MIT-licensed open-source skill/plugin for Claude Code, Codex, Gemini, Cursor, Windsurf, Cline, Copilot, and 30+…
Research signal · Discovered, not yet verified GitHub source ↗ for JuliusBrussee/cavemanDeepTutor HKUDS/DeepTutor is an Apache-2.0 open-source agent-native personalized learning platform from the HKU Data Science Lab that wraps LLM…
Research signal · Discovered, not yet verified GitHub source ↗ for HKUDS/DeepTutorcode-review-graph tirth8205/code-review-graph is an MIT-licensed open-source local-first code intelligence graph for MCP and CLI that builds a…
Research signal · Discovered, not yet verified GitHub source ↗ for tirth8205/code-review-graphopendatalab/MinerU is the Apache-2.0 (with a commercial-use threshold of 100M monthly active users or $20M monthly revenue, well above the vast…
Research signal · Discovered, not yet verified GitHub source ↗ for opendatalab/MinerUSource-backed candidates chosen for further research. This is an editorial watchlist, not a measured growth ranking; each card shows its current catalog score and retained verification stage.
LangChain is a Python/JavaScript agent-engineering framework for composing model calls, tools, retrieval, structured outputs, multi-agent flows, and provider integrations.
An AI-powered job search system layered on Anthropic's Claude Code CLI.
N8n is a Fair-code workflow automation platform with native AI capabilities.
LocalAI is a self-hosted OpenAI-compatible inference server for running LLMs, embeddings, image/audio generation, speech, rerankers, and vision-style workloads locally.
Dify is a low-code platform for building LLM apps: chatbots, agent workflows, RAG knowledge bases, prompt orchestration, deployment, and operational monitoring.
LobeHub is a web-based agent operations hub for creating specialized AI agents, scheduling them as always-on jobs, attaching knowledge bases, and reviewing their run reports.
The latest tools entering the research catalog. Filtered by recency and catalog score for triage; inclusion is not editorial verification.
OWASP Agent Memory Guard is an Apache-2.0 Python package for defending AI-agent memory against poisoning, tool abuse, privilege escalation, and excessive autonomy. Current...
FluxVLA is an Apache-2.0 engineering platform for vision-language-action work, covering data, model serving, and real-robot deployment workflows. Current GitHub metadata shows...
SecondSign Core is an Apache-2.0 Python project for runtime authorization around financial AI agents. The README describes structured intent, deterministic policy, human...
Council Lab is an Apache-2.0 local-first workspace for running a structured five-seat AI deliberation: four sequential agent viewpoints plus a final summarizer after a user...
Fara1.5-27B is Microsoft's MIT-licensed 27B multimodal computer-use model on Hugging Face. The model card says it observes browser screenshots and emits structured tool calls...
This arXiv paper introduces State Transition Pretraining for GUI agents: a multimodal model learns from visual before/after interface states by jointly predicting actions and...
The homepage separates reviewed decision surfaces from clearly labeled research signals. Search, filters, source links, risk labels, and all 2,349 ranked cards are available in the full catalog.
All ranked cards moved off the homepage.
Fresh and recently refreshed items in the public dataset.
Items with a practical verdict worth evaluating now.
Top evidence-linked useful signals.
RepoRadar is independently run. Tips help cover hosting, sources, and the time it takes to keep this useful. Never ranking. Never coverage. Never results.
V1 scoring prioritizes practical usefulness: usefulness 35%, novelty 18%, momentum 14%, maturity 10%, open-source/build quality 7%, evidence 6%, commercial/workflow potential 6%, setup ease 4%.
Gold is score ≥ 7.5 or the top 10% of eligible items with score ≥ 7.0. Silver starts at 6.5, Bronze at 5.25, and Low Signal catches low-value or low-signal items.
Gold requires Medium/High confidence, usable/official/cross-verified evidence or better, clear public descriptions, no excluded category, no weak or unclear AI relevance, and no fallback, triage-only, weak-evidence, low-signal, or low-transparency flags.
Risk remains a separate label and should not heavily reduce score unless severe. Transparency issues apply small penalties; hard exclusions still block publication.