Skip to content
{ ken ashe }
  • Building
  • Topics
  • Blog
  • Building
  • Topics
  • Blog
← Blog / Tags

// tag

multimodal-ai

11 posts tagged multimodal-ai.

  • OmniScientist argues that AI scientists need eyes, not just workflows Aug 14, 2026
  • AMIE’s video consult result is about perception, not replacement Aug 11, 2026
  • Video deep research agents need to look before they search Aug 5, 2026
  • MODUS brings any-to-any multimodal modeling to decoder-only systems Jul 29, 2026
  • Multimodal AI needs a plan for missing inputs Jul 28, 2026
  • MIRROR trains vision models by making each modality teach the others Jul 24, 2026
  • Sarcasm detection needs the mismatch, not just the meme Jul 20, 2026
  • Visual pretraining is a bet against text extraction Jul 13, 2026
  • Claude-real-video points to video as an adapter problem Jul 4, 2026
  • Speaker recognition is a better agent test than another chat demo Jul 3, 2026
  • Multimodal models still change answers when you shuffle the evidence Jun 25, 2026

Ken Ashe ·AI application builder ·CPA ·PMP

Building with AI in public. No hype, no doom. Receipts only.

hello@kenashe.ai

Social

  • LinkedIn
  • X
  • GitHub
  • RSS

Site

  • Building
  • Blog
  • Newsroom
  • Media Kit
  • Lucky Domains

Legal

  • Privacy
  • Terms
  • Disclosure

© 2026 Ken Ashe ·Built with AI in public