Essential cookies keep Instavar working. Optional analytics help us understand how the site is used and link your first-visit source to your Studio account after sign-in, for up to 180 days. Cookie Policy

Manage Cookie Preferences

Service reliability telemetry, including Sentry error monitoring and Vercel Speed Insights, stays enabled so we can secure the product and diagnose failures.

Skip to content
ResearchPlaybooksToolsOpen sourceCase studiesStudio
  1. Home
  2. Blog

Singapore Office - Antiphishing Pte. Ltd.

JTC LaunchPad @ one-north

67 Ayer Rajah Crescent, #02-14

Singapore 139950

© 2026 Instavar. All rights reserved.

Find the angle that travels.

ResearchPlaybooksToolsOpen sourceCase studiesStudioAboutCareers|PrivacyTermsCookiesAI PolicyRights & ConsentReport Abuse

UK Corporate Presence - Litiga Ltd

Registered in England and Wales

Company number 11610573

Registered office: 128 City Road

London, England EC1V 2NX

Incorporated 8 October 2018

Instavar Blog

Make better content decisions.

Practical lessons on video, distribution, and measurement. Start with the question you need to answer, then go deeper when it helps.

Latest notes.

  • Running OpenAI Privacy Filter on an M2 MacBook Pro - 52-Case Benchmark

    Can OpenAI's new privacy-filter token classifier redact secrets and PII on a 16 GB M2 MacBook before they reach a cloud LLM? We ran 52 test cases covering API keys, PEM blocks, JWTs, names across six cultures, addresses, and decoy prose. Load time, latency, hit rates, and the AWS-key miss that matters.

  • How Open-Source TTS Architectures Differ - And What It Means for Fine-Tuning (2026)

    A practitioner's comparison of 6 TTS architectures (Voxtral, Qwen3-TTS, IndexTTS2, Chatterbox, Fish Speech, CosyVoice 3) covering codecs, LoRA compatibility, data pipelines, and license traps. Based on first-party adaptation work on the same FEMALE_01 corpus slice.

Build an AI YouTube Shorts Pipeline - Remotion + TTS + Automated Publishing

Architecture for an automated Shorts pipeline that survived 136 render cycles. Covers Remotion composition, TTS integration, Inngest orchestration, and multi-platform publishing.

  • DeepSeek OCR-2 in Production - What the Benchmarks Don't Tell You

    Production notes from running DeepSeek OCR-2 on 50 scanned pages. Covers markdown output quality, blank detection strength, and where it falls behind GLM and Qianfan. First-party CER data included.

  • F5-TTS Fine-Tuning Guide 2026 - Colab, Quality, VRAM, and Voice Cloning

    F5-TTS fine-tuning guide for custom voice cloning in 2026. Covers Colab-style setup, dataset preparation, VRAM, quality review, ease of fine-tuning, and comparison with Qwen3-TTS, VoxCPM, and CosyVoice.

  • Hunyuan OCR vs FireRed OCR - Which Handles Your Documents Better?

    Side-by-side comparison of Hunyuan OCR and FireRed OCR on 50 scanned pages across 7 document types. Includes CER scores, latency, and routing recommendations from our first-party benchmark.

  • Best OCR for Scanned PDFs - 5 Models Tested on 50 Scanned Pages

    We tested 5 OCR models on 50 scan-heavy pages across 7 document types. Here's which model to use for clean scans, degraded scans, tables, formulas, and mixed documents.

  • Which TTS Model Should You Use? A Decision Tree (2026)

    Open-source TTS model comparison for 2026 across Qwen3-TTS, CosyVoice, F5-TTS, Fish Speech, Kokoro, Supertonic, VoxCPM, IndexTTS2, Chatterbox, and Higgs Audio. Compare fine-tuning, VRAM, inference speed, latency, edge deployment, and voice quality.

  • YouTube Shorts for AI-Generated Content - Rules, Monetization, and What Gets Flagged

    YouTube's rules for AI-generated Shorts in 2026. Covers disclosure requirements, Content ID restrictions, monetization eligibility, and what actually gets flagged.

  • YouTube Shorts Retention Curve - Read It, Fix It, Automate It

    How to read YouTube Shorts retention curves, diagnose the five common failure patterns, and build an automated feedback loop from analytics to production.

  • Best Open-Source TTS Models for Production in 2026

    A practitioner's comparison of open-source TTS models and Instavar's voice-adaptation repository programme on a 24GB GPU. Covers fine-tuning paths, checkpoint selection, runtime qualification, failure modes, and evidence boundaries.

  • CosyVoice Fine-Tuning Guide - LoRA, Data Requirements, and Voice Quality

    CosyVoice fine-tuning guide for voice cloning on a consumer GPU. Covers LoRA vs full SFT, data requirements, VRAM, voice quality, the epoch 12 rerun, 9 known pitfalls, and the PEFT companion repo.

  • Previous

    Page 3 of 11

    Next