Project Yoru & Fami

Stealth research on next-generation image generation and speech/audio foundation systems

Project Yoru and Project Fami are ongoing stealth research efforts exploring novel architectures and scaling infrastructures for generative vision and multimodal audio synthesis.

Stealth Research Data & Architecture Exploration

Research Focus Areas

🌌 Project Yoru (Vision & Image Generation)

  • High-Fidelity Flow Map Formulation: Exploring continuous and decoupled flow matching formulations for accelerated, high-fidelity text-to-image synthesis.
  • Architectural Exploration: Designing scale-wise transformer backbones and representation autoencoders (RAE) to eliminate traditional VAE artifacts.
  • Dataset Infrastructure: Developing automated high-resolution image filtering, aesthetic scoring, and synthetic captioning pipelines.

🎙️ Project Fami (Audio & Speech Synthesis)

  • Generative Audio Framework: Researching continuous flow matching and diffusion-based acoustic decoders for natural speech and audio synthesis.
  • Data Engineering & Alignment: Constructing clean, multi-speaker conversational speech corpora with fine-grained prosody and acoustic feature extraction.