Explore / GitHub Agent Apps Campaign

GitHub Agent Apps Campaign

DeepQA runs its QA engine against agent apps that are trending on GitHub: workflow builders, chat UIs, coding agents, research agents and observability tools. Open-source apps are cloned and booted from their repository, hosted apps are tested in place, and every report is published here.

A team that wants its report taken down can ask us and we will remove it.

56
apps
56
public reports
631
scenarios
562
scenarios passed
28
issues confirmed
54
re-verified live
2k
screenshots
9.6k
browser actions
9.1k
model calls
57.2M
tokens
8.5 h
testing time

56 apps (showing 1 to 12)

subtitle-translator in the browser during the run
Subtitle Translatorsubtitle-translatorrockbenben/subtitle-translatorLLM translation toolSandbox, from sourcerun by DeepQA Team

Open-source batch subtitle translator with many LLM and translation providers, multi-language output and provider settings. Run from its repository in the DeepQA sandbox.

Run #1succeeded12 scenarios, 12 pass, 0 fail
0 issues
m3e-canvas in the browser during the run
M3E Canvasm3e-canvaslnkiai/m3e-canvasDesign canvas for agent promptsSandbox, from sourcerun by DeepQA Team

Open-source design canvas that turns Material 3 Expressive screens into prompts for coding agents: screens, layers, themes, preview and prompt export. Run from its repository in the DeepQA sandbox.

Run #1succeeded12 scenarios, 12 pass, 0 fail
0 issues
ChatChat in the browser during the run
Chat ChatChatChatokisdev/ChatChatChat UISandbox, from sourcerun by DeepQA Team

Open-source unified chat and search platform across model providers with sessions, system prompts and provider settings. Run from its repository in the DeepQA sandbox.

Run #1succeeded12 scenarios, 10 pass, 1 fail, 1 blocked
0 issues
OpenClaw-bot-review in the browser during the run
OpenClaw Bot ReviewOpenClaw-bot-reviewxmanrui/OpenClaw-bot-reviewAgent fleet dashboardSandbox, from sourcerun by DeepQA Team

Open-source dashboard for a fleet of OpenClaw agents: bot wall, models, sessions, statistics and an alert rule center. Run from its repository in the DeepQA sandbox.

Run #1succeeded12 scenarios, 8 pass, 4 fail
3 issues21
claudecodeui in the browser during the run
Claude Code UIclaudecodeuisiteboon/claudecodeuiCoding agent web UISandbox, from sourcerun by DeepQA Team

Open-source web and mobile UI for coding agent CLIs: setup wizard, project workspaces, chat, shell, files and source control views. Run from its repository in the DeepQA sandbox.

Run #1succeeded12 scenarios, 10 pass, 1 fail, 1 blocked
0 issues
crit.md in the browser during the run
Critcrit.mdOpen app ↗tomasz-tomczyk/critHosted appAgent feedback toolHosted, in placerun by DeepQA Team

Open-source review tool for coding agents: comment on lines, reply in threads and turn the review into an agent prompt. The public site carries an interactive review demo. Tested in place without signing in.

Run #1succeeded12 scenarios, 11 pass, 1 fail
1 issue1
ggml-org-gguf-my-repo.hf.space in the browser during the run
GGUF My Repoggml-org-gguf-my-repo.hf.spaceOpen app ↗ggml-org/llama.cppHosted appModel conversion toolHosted, in placerun by DeepQA Team

Hugging Face Space by the llama.cpp organization that converts and quantizes a model repository to GGUF. Tested in place without signing in, so the conversion itself was not run.

Run #1succeeded7 scenarios, 7 pass, 0 fail
0 issues
qwen-qwen3-demo.hf.space in the browser during the run
Qwen3 demoqwen-qwen3-demo.hf.spaceOpen app ↗QwenLM/Qwen3Hosted appLLM chat playgroundHosted, in placerun by DeepQA Team

Official Hugging Face Space of Qwen3 by the Qwen team: chat with a thinking mode toggle, thinking budget and conversation history. Tested in place.

Run #1succeeded12 scenarios, 12 pass, 0 fail
0 issues
agent-flow in the browser during the run
Agent Flowagent-flowpatoles/agent-flowAgent run visualizerSandbox, from sourcerun by DeepQA Team

Open-source real-time visualizer for coding agent sessions: agent graph on a canvas, timeline, transcript, file attention and cost overlay. Run from its repository in the DeepQA sandbox.

Run #1succeeded12 scenarios, 11 pass, 1 fail
0 issues
neo.u14.app in the browser during the run
Neo Chatneo.u14.appOpen app ↗u14app/neo-chatHosted appChat workspaceHosted, in placerun by DeepQA Team

Open-source local-first AI chat workspace with assistants, skills, plugins, knowledge bases, workspaces and MCP servers. Tested in place on its hosted instance.

Run #1succeeded12 scenarios, 10 pass, 2 fail
1 issue1
logocreator in the browser during the run
Logo CreatorlogocreatorNutlope/logocreatorAI logo generatorSandbox, from sourcerun by DeepQA Team

Open-source AI logo generator with styles, colors, brand kits and a bring-your-own-key flow. Run from its repository in the DeepQA sandbox.

Run #1succeeded12 scenarios, 11 pass, 1 fail
1 issue1
gpt_image_playground in the browser during the run
GPT Image Playgroundgpt_image_playgroundCookSleep/gpt_image_playgroundImage generation playgroundSandbox, from sourcerun by DeepQA Team

Open-source bring-your-own-key playground for GPT image models with a gallery, favorite folders, search and API settings. Run from its repository in the DeepQA sandbox.

Run #1succeeded12 scenarios, 11 pass, 1 fail
0 issues

Put an agent team on your next pull request.

Connect a repo, dispatch a Run, and read an audited, evidence-backed report the same day.