K-001 / Studio Overview BOSTON · LOS ANGELES · SHANGHAI

An AI-native studio
for moving images.
Built for production.

Frame-aware
Catalog-scale
Shipped to App Store and Play Store

We build AI for streaming platforms, content libraries, and consumer media apps. Scene tagging, clipping, dubbing, captions, promo cuts, on-device inference, signed provenance. Catalog work that moves the numbers paying for the catalog: CPMs, fill rate, watch-time, trial conversion. Apps that actually ship.

K-002 / Capabilities 04 CORE · 09 EXTENDED

A studio with a complete toolset.

Research, ML engineering, backend, mobile, and design under one roof. Thirteen practices in total. You can engage us on any single one or string several together. Each engagement comes with a deploy URL, a build number, or both.

/ 01

Scene Intelligence APIs

Frame-level object, scene, and dialogue understanding for OTT catalogs. Temporal-spatial metadata, rights-aware exports, and downstream feeds for content ops, ad sales, and search. The data layer that ad sales has been asking for.

  • SAM3 · YOLO26
  • Whisper · Clova
  • Multimodal indexing
Open practice
/ 02

Clip & Highlight Engines

Bulk short-form generation for streaming catalogs. Drives social tune-in to the catalog and lifts SVOD trial conversion. Tuned for Korean, Japanese, Chinese-language, and English long-form. Editors keep the final say on every clip.

  • Highlight ranking
  • 9:16 reframe
  • Subtitle alignment
Open practice
/ 03

Mobile AI Apps · iOS / Android

Native media apps with on-device inference: CoreML, MLX, MediaPipe, ONNX. Recording, transcription, replacement, super-resolution. Designed offline-first and shipped to both stores. The cloud GPU bill stays small.

  • Swift · CoreML
  • Kotlin · NNAPI
  • React Native bridges
Open practice
/ 04

Custom Pipelines & R&D

When off-the-shelf doesn’t fit. Bespoke inference pipelines, fine-tunes, evaluation harnesses, internal tooling. We staff the seam between research and production so a line in a paper turns into a line in your delivery spec.

  • Diffusion · finetune
  • GPU orchestration
  • Eval harness
Open practice
/ Extended Practices · 09 Modules that plug into the four practices above, or ship on their own. Each one maps to a line item your CFO understands: ad yield, fill rate, watch-time, marketing throughput, compliance, retention.
/ 05

Dubbing & Lip-Sync

Voice-faithful dubs at season scale. Lip retiming, M&E stem isolation, QC that runs alongside your existing localization vendor. 30+ languages, which is the cheapest way to grow addressable audience for a back-catalog title.

  • ElevenLabs · Papercup
  • wav2lip · Diff2Lip
  • IMSC · TTML2
Open practice
/ 06

Subtitles, SDH & Audio Description

Captions, SDH, and audio description that meet ADA Title II, the EAA, and platform spec. Whisper-class ASR, frame-aligned timing, operator review only on exceptions. Unlocks the audience segments your accessibility lead has been flagging.

  • Whisper · Speechmatics
  • IMSC Rosetta · WebVTT
  • WCAG 2.1 AA
Open practice
/ 07

Trailer & Promo Auto-Cut

Promo cutting that reads pacing, beats, and music cues. Drafts trailers, teasers, key-art motion, and per-territory variants for editor sign-off. Cuts marketing turnaround and opens promo coverage on the long tail your team can’t staff.

  • Pegasus · Marengo
  • MusicLM · beat-track
  • Per-territory variants
Open practice
/ 08

Live & Sports Highlights

HLS, RTMP, or SRT in. Vertical highlights out, event-triggered (goal, wicket, breakaway), multilingual, with auto-publish to social and OTT. Built to drive same-day tune-in and post-event sponsor inventory.

  • Whisper-Live · NeMo
  • OpenCV · YOLO26
  • Event-triggered
Open practice
/ 09

Contextual Ad Metadata

Mood, scene, and brand-safety signals packaged for programmatic CTV. Lifts CPMs on premium inventory, opens fill from blue-chip brands that won’t buy uncertified context, and stays PII-free. IRIS_ID-compatible. Ships into Samba, GumGum, Wurl, or your own SSP.

  • IRIS_ID · IAB v3
  • BERT · CLIP
  • GARM · brand-safety
Open practice
/ 10

Catalog QC & Delivery

Lip-sync checks, subtitle-timing validation, spoken-language verification, loudness and burnt-in conflict detection. Sits next to Vantage or Pulsar in your QC pipeline rather than replacing them. Stops the rejection-and-redeliver cycle that bleeds margin.

  • Telestream · Venera
  • EBU R128 · IMSC
  • pyannote · lipnet
Open practice
/ 11

Provenance & C2PA

C2PA 2.1 manifests at ingest, signed edit chains through the DAM, SynthID and AudioSeal on AI-touched assets. Built around EU AI Act Article 50 (in force August 2026) so a compliance miss never delays a delivery.

  • C2PA 2.1 · ISO 22144
  • SynthID · AudioSeal
  • Truepic · Adobe CCI
Open practice
/ 12

Recsys Embeddings

Scene-essence vectors for “more like this” rails, mood search, and editorial collections. Exports to Personalize, FAISS, or pgvector. Drives session length, completion rate, and SVOD retention without forcing a recsys re-platform.

  • CLIP · SigLIP
  • FAISS · pgvector
  • Vionlabs-style
Open practice
/ 13

Roto & Object Removal

Text-prompt mask and track for compliance blurs, logo swaps, talent scrubs, and VFX prep. Built to run across thousands of episodes with masks that survive cuts. Also the substrate for in-scene product placement when sponsorship comes calling.

  • SAM3 · YOLO26
  • Grounding DINO
  • ByteTrack · BoT-SORT
Open practice
15yr
Inside OTT & streaming
200hr
Catalog throughput / day
2–4×
CPM lift on contextual inventory
iOS · Android
Native, on-device, signed
K-003 / Featured Work NOW SHIPPING

Things we’ve made that ship.

Two flagship products, plus client work under NDA across U.S. and Asian streaming. Each one runs against real catalogs and is paid for by a real P&L line.

WORK / 01 · KLADON CLIPS2026

Kladon Clips — catalog-scale short-form for OTT.

Pulls the high-engagement moments out of long-form OTT libraries. Platform-ready TikTok, Reels, and Shorts at ~200hr/day, with editor sign-off. Drives social tune-in on the catalog you’ve already paid for.

Read the case
WORK / 02 · SCENE INTELLIGENCE ENGINE2026

Scene Intelligence Engine — every object, every frame.

API layer for VPP partners, ad-tech, and post-production. Object taxonomy with temporal-spatial metadata, plus a replacement-ready downstream pipeline. The data underneath premium contextual inventory and direct-sold scene-level packages.

Read the case
K-004 / Operating Principles 05

We’re AI-native. We work in frames rather than slides, ship code rather than decks, and treat models as interchangeable components.

/ 01
Models are components.
No single model wins everything. We pick the right tool for the task (open weights, hosted API, or a fine-tune) and swap it without drama when something better lands. That happens often.
/ 02
Editorial sign-off, not autopilot.
Brand-grade media still needs a human in the loop. We build editor-assist tools where automation does the bulk and the editor stays in command. Even a strong model isn’t a perfect one.
/ 03
Deploy on day one.
Every engagement starts with a deploy URL or a TestFlight build. Progress is measured in build numbers and inference logs, not Jira tickets.
/ 04
Rights, watermarks, provenance.
Media AI without C2PA, music licensing, and rights metadata is a lawsuit waiting to happen. We build compliance in from day one: EU AI Act Article 50, ASCAP, BMI, JASRAC, KOMCA, MCSC royalties, per-output watermarks.
/ 05
Money has to follow the work.
Every engagement has to land in a P&L line: ad CPMs, fill rate, ARPU, watch-time, retention, marketing throughput, compliance cost avoided. If we can’t name the line, we don’t take the work.
K-005 / Process 04 PHASES

From signal to ship.

Every engagement runs the same loop. The team that writes the eval harness is the team that ships the iOS binary, which is how we keep the gap between research and production tight.

01
Diagnose
Two-week scoping sprint. We watch your content, audit the stack, talk to the editors and ad sales people who actually use the data. Output: a written brief and a target P&L line.
Wk 01–02
02
Design
Architecture, data flow, model selection, evaluation criteria. The eval harness is written before the inference pipeline, not after.
Wk 02–04
03
Deploy
First production output by week 4. Iteration is measured against editor sign-off rate, CPM lift, fill, or whatever the target line was at week 1.
Wk 04–10
04
Operate
Monthly retainer or a clean hand-off package. The field moves week to week, so we keep your models, prompts, and weights current.
Wk 10+

Building something frame-aware?

We take a small number of engagements each quarter. Mostly OTT operators, content studios, post houses, and product teams shipping consumer media apps.

If you have a long-form catalog, an editorial bottleneck, an unsold AVOD pod, or an AI app that needs to be on the store this quarter, write us. Replies inside two business days.