Recon area

Candidates the watcher flagged, plus anything you drop in. This is the inbox of the Recon lane: vet each one on a fast, cheap pass, then promote the winners into the model registry or reject them. Run python tools/watch_models.py (or the scheduled task) to refresh it. See the watcher spec.

Category
Status
Why
videonewhyped

Kandinsky 5.0 Video Lite

Fast 2B draft-tier video. Bake off against HunyuanVideo 1.5 on a cast-rich song.

Flagged: 2026-07-20
source →
controltestingnew

Wan-VACE

Core of the control rig (pose/depth/ref). Test a directed hero shot end to end.

Flagged: 2026-07-20
source →
ttsnewhyped

OmniVoice

600+ language zero-shot cloning; the multilingual wildcard for the dub lane.

Flagged: 2026-07-20
source →
lipsyncnewhyped

Wan-S2V

Audio-driven cinematic avatar on the Wan stack; paper reports it beats HunyuanVideo-Avatar and OmniHuman. Slots into the existing Wan setup, test next.

Flagged: 2026-07-21
source →
ttsnewhyped

Fish Audio S2 / S2-Pro

Tops independent TTS benchmarks (paralinguistics), sub-100ms, 80+ langs. LICENCE non-commercial (research); commercial needs paid licence. Quality ceiling only.

Flagged: 2026-07-21
source →
llmnewhyped

Kimi K2.6

Top open coding model (~1T MoE, modified MIT); needs serious rented hardware. For big-GPU scripting runs.

Flagged: 2026-07-21
source →
videonewhyped

Wan 2.7

UNVERIFIED: reported April 2026 with first/last-frame control and 9-grid input, but open-weight release NOT confirmed (Wan 2.5 was API-only). Confirm weights are actually open before adopting.

Flagged: 2026-07-21
source →
imagenewhyped

Cosmos3-Super-Text2Image

UNVERIFIED: claimed open T2I leader (~1219 Elo, 65B) from a single low-quality source. Needs primary confirmation before trusting.

Flagged: 2026-07-21
source →

Where candidates come from

Three kinds of signal, used together. The watcher automates the popularity scan; these add quality and practitioner judgement. The funnel stays: discover, rank, then vet on your own footage.

Leaderboards

Rank quality and preference
LMArena llm, image
Human-preference Elo for LLMs, with image and vision arenas.
Artificial Analysis llm, video, tts
Quality vs speed vs price across LLMs, plus video and speech arenas.
LiveBench llm
Contamination-resistant LLM scoring, refreshed over time.
VBench video
The standard open benchmark leaderboard for video generation.
imgsys image
Blind A/B image-model arena (by fal).
TTS Arena tts
Blind A/B voice-preference ranking for text-to-speech.
Open ASR Leaderboard stt
The authoritative WER ranking for speech-to-text.
MTEB embeddings
The canonical embedding-model benchmark (where the Qwen3-Embedding pick came from).

Trusted channels

Practitioner signal, ahead of the boards
Theoretically Media video, film
Tim Simmons tests AI video/image/audio tools honestly in real production; no tool-of-the-week hype.
Curious Refuge video, film
AI filmmaking school and news; workflows, shorts and the weekly AI-film roundup.
Aitrepreneur image, video, tts, llm
Broad, hands-on open/local model coverage: Stable Diffusion, LoRA training, voice cloning, local LLMs, including uncensored and NSFW open models. Not just one category.

Community and discovery

Raw new releases
Hugging Face Trending all
What the recon watcher already taps; popularity, not quality.
Papers with Code all
State-of-the-art tables per task, with linked code.
r/LocalLLaMA llm
Fast community signal on new open models and quantised builds.
r/StableDiffusion image, video
Community signal on image/video models, LoRAs and ComfyUI workflows.