Question 1

Which is better: LamBench or last30days-skill?

Accepted Answer

Based on our expert panel, last30days-skill has a stronger verdict with a 75% Ship rate. LamBench received a panel verdict of Mixed and last30days-skill received Ship.

Question 2

Is LamBench free?

Accepted Answer

LamBench pricing: Free / Open Source

Question 3

Is last30days-skill free?

Accepted Answer

last30days-skill pricing: Free / Open Source (API keys needed for full features)

Question 4

What do experts say about LamBench vs last30days-skill?

Accepted Answer

LamBench: LamBench is a benchmark of 120 fresh lambda calculus programming questions designed by Victor Taelin (creator of the HVM runtime) to test genuine AI reasoning capabilities rather than pattern-matched performance on contaminated datasets. Questions range from implementing basic operations like addition for λ-encoded natural numbers to deriving generic folds for arbitrary data types.

The benchmark measures both accuracy (percentage of 120 tasks solved correctly) and speed (average solution time). Current top performers include GPT-5.4 at 91.7% accuracy, Anthropic's Opus 4.6 at 90.0%, and GPT-5.3-Codex at 89.2%. Lower-tier models bottom out at 28-58% accuracy — revealing significant gaps in symbolic reasoning capability that other benchmarks obscure.

Taelin released LamBench in direct response to community requests for a benchmark resistant to training data contamination. Lambda calculus is a clean, closed formal system — ideal for testing reasoning because memorizing examples provides minimal advantage over actually understanding the abstractions. last30days-skill: last30days-skill is an AI agent skill that aggregates, deduplicates, and synthesizes recent discussions about any topic from Reddit, X/Twitter, YouTube, Hacker News, Polymarket, Bluesky, TikTok, and Instagram simultaneously. The core value proposition: instead of manually searching eight platforms and stitching together what people are actually saying, you ask once and get a grounded summary with citations ranked by engagement and cross-platform convergence.

The ranking system is unusually sophisticated for a community project—it combines text similarity, engagement velocity, source authority, and cross-platform convergence detection (penalizing topics that only appear on one platform). For prediction markets, it evaluates topics as outcomes within broader events rather than naive title matching. A handle resolution feature identifies X/Twitter accounts from natural language names alone. Zero configuration is needed for Reddit, HN, and Polymarket; unlocking other sources requires API keys from ScrapeCreators and Exa.

The project reached 18k stars in its first week, largely driven by prompt researchers discovering it surfaces "what actually works" for tools like ChatGPT or Midjourney. Results auto-save to ~/Documents/Last30Days/ by default, and a watchlist mode supports scheduled topic monitoring with an external cron scheduler.

LamBench vs last30days-skill

LamBench

last30days-skill

Bookmarks