Tested,
not hyped.

Nowness is an autonomous AI lab that runs itself — on local models, on one machine, around the clock. It hunts the frontier of AI research, runs the new tools for real to prove what works, turns the winners into usable use-cases, and invents its own.

What the lab does — on its own, non-stop
01 · Discover

Hunts the frontier

Finds the newest AI research and tools the moment they appear.

02 · Prove

Runs it for real

Clones, installs, and executes each one in a locked-down sandbox — truth, not README claims.

03 · Translate

Research → use‑cases

Turns what actually works into real, usable use-cases.

04 · Invent

Builds new tech

Combines what it's learned into its own working prototypes — and proves they run.

0%

One thing it proves: 1,043 AI repos it actually ran, and a third don't work.
Everyone judges AI by the demo. Nowness runs the code — and only surfaces what's real.

Try it

Send Nowness a repo.

Paste any public GitHub repo and your email. Nowness clones it, installs it, and actually runs it in a locked-down sandbox — you watch the whole test happen live, right here.

Here's exactly what lands in your inbox:

Does it really install & run An honest verdict tier The real evidence — tests passed, demo output A screenshot of it running

1,671 repos tested by the lab so far

The daily pick · under the radar

Today's verified pick.

Every day Nowness features ONE repo from its verified winners — ranked purely by real execution evidence (tests that passed, installs that worked, demos that ran), never by stars, and never an obvious big name. A fresh verified gem, daily.

run‑verified · sandbox
★ DAILY PICK · 24 Jul 2026 ✓ production-ready MCP server

EvoAgent

EvoAgent is an asynchronous, modular LLM agent framework built around a ReAct loop for tool use, memory evolution, and multi-agent orchestration.

692tests passed
~7★github stars
18 Julverdict earned
Why it's today's pick — exactly

EvoAgent is an asynchronous, modular framework designed for multi-agent orchestration and tool use. It employs a ReAct loop for memory evolution and a deterministic-first code retrieval stack. The system includes an MCP client and crash-safe resume capabilities with tracing for better observability. Laboratory tests confirmed the framework's functionality through a comprehensive test suite and successful demonstration runs.

The project earns its spotlight by addressing the lack of reliability and state persistence in current agent frameworks. It enables autonomous software engineering and complex workflow automation by allowing for task decomposition and parallel execution. By providing a reproducible evaluation harness, it offers a stable path for managing long-running tasks and complex tool interactions.

Live

What the lab is testing.

Nowness tests continuously — trending repos, papers, and whatever you send. This is live from the sandbox.

Verified finds

Real repos. Real runs.

Every card below was actually executed by the lab — under-the-radar repos that installed clean and did what they claim, verified in the sandbox, not guessed from the README. From 1,671 repos tested so far.

IRIS (Intelligent Research Insight System)

IRIS is an automated research and report generation system built on an Agentic Workflow using LangGraph.

Insight Installed cleanly on the first try.

github.com/ttguy0707/IRIS ↗

Stabilize

Stabilize is a durable, queue-based state machine and workflow engine for Python that orchestrates Directed Acyclic Graphs (DAGs).

Insight The project has a comprehensive structure, multiple examples, and a clear API.

github.com/rodmena-limited/stabilize ↗

ArcadeDB

ArcadeDB is a high-performance Multi-Model Database Management System that supports SQL, Cypher, Gremlin, HTTP/JSON, MongoDB, and Redis queries within a single engine.

Insight The project has a comprehensive structure, multiple release tags, and a clear multi-model implementation.

github.com/ArcadeData/arcadedb ↗

SparseIndex

A Rust library for high-performance sparse vector indexing, retrieval, and storage.

Insight Installed cleanly on the first try; its own test suite ran — 109 tests passed.

github.com/myscale/sparse-index ↗

Git Graph for VS Code

A Visual Studio Code extension that provides a graphical representation of Git repository history.

Insight Its own test suite ran — 1,266 tests passed.

github.com/mhutchie/vscode-git-graph ↗

Typer

Typer is a Python library for building command-line interfaces (CLIs) that leverages Python type hints for automatic validation and completion.

Insight The demo actually ran and produced real output; the library imports without errors.

github.com/tiangolo/typer ↗
Browse the full database of verified finds →

Stop guessing. Send a repo.

Nowness will tell you whether that trending repo actually works — with the evidence.