VerifyFirst is an evidence-backed scam investigation agent built on TrueForge. Users can submit a suspicious message, URL, phone number, email, and claimed organization in one investigation. VerifyFirst independently checks identity, sender reputation, domains, URLs, email infrastructure, trusted sources, and suspicious behavioral signals, then correlates the evidence into a clear risk assessment and safe next action. TrueForge is the execution harness, not a wrapper: VerifyFirst uses MCP tools, native subagents, sandbox execution, durable sessions, execution traces, and human approval before exporting sensitive evidence. Key capabilities: - Multi-signal scam and impersonation investigation - Live phone reputation through IPQualityScore - URL/domain and organization verification - Email DNS, MX, SPF, DMARC, RDAP, and lookalike checks - Prompt-injection resistance for untrusted content - Evidence-backed findings with uncertainty preserved - Native TrueForge approval-gated report export - Human-readable and machine-readable evidence reports - Reproducible local and Docker setup - Automated tests and evals The goal is simple: do not ask an LLM whether something “looks like a scam.” Independently verify the evidence before a person sends money, credentials, or sensitive information.
Built at The Agent Harness Hackathon
VerifyFirst
VerifyFirst is an evidence-backed scam investigation agent built on TrueForge. Users can submit a suspicious message, URL, phone number, email, and claimed organization in one investigation. VerifyFirst independently checks identity, sender reputation, domains, URLs, email infrastructure, trusted sources, and suspicious behavioral signals, then correlates the evidence into a clear risk assessment and safe next action. TrueForge is the execution harness, not a wrapper: VerifyFirst uses MCP tools, native subagents, sandbox execution, durable sessions, execution traces, and human approval before exporting sensitive evidence. Key capabilities: - Multi-signal scam and impersonation investigation - Live phone reputation through IPQualityScore - URL/domain and organization verification - Email DNS, MX, SPF, DMARC, RDAP, and lookalike checks - Prompt-injection resistance for untrusted content - Evidence-backed findings with uncertainty preserved - Native TrueForge approval-gated report export - Human-readable and machine-readable evidence reports - Reproducible local and Docker setup - Automated tests and evals The goal is simple: do not ask an LLM whether something “looks like a scam.” Independently verify the evidence before a person sends money, credentials, or sensitive information.
Keep exploring what builders shipped.
Sports Whisperer
SportsWhisperer
With AI becoming centre stage - Content and Entertainment will be the king. Sport has the highest amount of spend as industry and creates massive economic drive, so we have built an All in One - Sports Whisperer for all major sports Crickets, NFL, Soccer, Basketball, Baseball, etc where the novice and experienced players can interact with the favorite games and players !!! More details captured - https://docs.google.com/presentation/d/1_q0SDTvutQP9kd2SLdDXYYkc-h9QVYwj0vdih1Cuqzc/edit?slide=id.gcb9a0b074_1_0#slide=id.gcb9a0b074_1_0
AgentInvariant
AgentInvariant
AgentInvariant is a behavioral safety evaluator for tool-using AI agents. Agent workflows are probabilistic: two requests with the same meaning can produce materially different external actions. AgentInvariant uses metamorphic testing to determine whether critical operational behavior remains consistent across meaning-equivalent inputs. For each input variant, AgentInvariant starts a fresh OpenAI conversation and an isolated SQLite database. The target agent uses real tools to check coverage, retrieve authorization requirements, record business approval, and submit a synthetic prior-authorization request. AgentInvariant records the complete structured tool trace and evaluates it using deterministic Python invariants—not an LLM judge. It verifies that exactly one submission occurs and that matching coverage and business approval are successfully recorded before submission. The evaluator is exposed through a Streamable HTTP MCP server and invoked by a TrueForge agent using the `evaluate_agent` tool. Results include per-variant PASS or BLOCKED decisions, ordered traces, exact invariant violations, compliance rate, and the shortest failing counterexample. The demo compares an explicitly labeled unsafe negative control, which AgentInvariant correctly blocks, with a hardened candidate that follows the required workflow. All healthcare data is synthetic. This project is an administrative workflow reliability demonstration and is not clinical decision support.
HackerSquad project
Robo Harness
Hardess for robo

