AI & ML interests

Building a globally standardized, AI-driven digital health and long-term care platform for aging societies — FHIR-first, with integrated health data and intelligent decision support for personalized preventive medicine. AI accelerates digitization; it isn't the product.

Recent Activity

gloomcheng  updated a Space 1 day ago
weemed/README
gloomcheng  updated a model 1 day ago
weemed/README
gloomcheng  updated a model 1 day ago
weemed/IlhaEmbed
View all activity

Organization Card

WeeMed AI

Building a globally standardized, AI-driven digital health and long-term care platform for aging societies.

FHIR-first. Integrated health data, intelligent decision support, personalized preventive medicine — deployed in real clinics, health-screening centers, and community long-term care sites in Taiwan.

AI accelerates digitization; it isn't the product.


🇹🇼 Taiwan Sovereign Medical AI Stack

We ship production-grade models derived from real-world clinical and health checkup deployment problems in Taiwan, publishing open benchmarks and transparent data provenance.

🧠 IlhaEmbed (v2.0)

Traditional Chinese Clinical & Medical Terminology Embedding Model (38.5 MB INT8 ONNX / 384-dim)

  • Domain-Adapted for Taiwan Clinical Writing: Reads Taiwanese hospital shorthand ( → 低劑量胸部電腦斷層), checkup note acronyms ( → 傷寒篩檢糞便檢體), clinical slang ( → 帶狀皰疹), and Taigi colloquialisms ( → 中風).
  • MODA & NAER 13-Set Retrained: Fully retrained with the Ministry of Digital Affairs (MODA) and National Academy for Educational Research (NAER) 111,386 medical concept taxonomy.
  • Context-Conditioned Disambiguation: Dynamically disambiguates pure-Latin acronyms (, , Bun is a fast JavaScript runtime, package manager, bundler, and test runner. (1.3.14+0d9b296af)

Usage: bun [...flags] [...args]

Commands: run ./my-script.ts Execute a file with Bun lint Run a package.json script test Run unit tests with Bun x vite Execute a package binary (CLI), installing if needed (bunx) repl Start a REPL session with Bun exec Run a shell script directly with Bun

install Install dependencies for a package.json (bun i) add @evan/duckdb Add a dependency to package.json (bun a) remove redux Remove a dependency from package.json (bun rm) update @zarfjs/zarf Update outdated dependencies audit Check installed packages for vulnerabilities outdated Display latest versions of outdated dependencies link [] Register or link a local npm package unlink Unregister a local npm package publish Publish a package to the npm registry patch Prepare a package for patching pm Additional package management utilities info zod Display package metadata from the registry why tailwindcss Explain why a package is installed

build ./a.ts ./b.jsx Bundle TypeScript & JavaScript into a single file

init Start an empty Bun project from a built-in template create vite Create a new project from a template (bun c) upgrade Upgrade to latest version of Bun. feedback ./file1 ./file2 Provide feedback to the Bun team.

--help Print help text for command.

Learn more about Bun: https://bun.com/docs Join our Discord community: https://bun.com/discord, ) and polysemous terms () using operational field context.

  • Pure CPU On-Premise Native: Zero GPU, zero cloud, zero API keys required. Patient data stays safely inside hospital premises.

🎙️ Breeze-ASR-26-edge & Taiwanese-Tailo-ASR

Taiwanese Hokkien (台語) + Mandarin Clinical Speech Recognition Stack

  • Quantized for edge devices in CTranslate2 () and ONNX runtimes.
  • Multi-format output support (, , , ).
  • Derived from MediaTek Research's Breeze-ASR-26 & OpenAI Whisper-large-v3, fine-tuned on SuíSiann 2.0, TAT_MOE, and real meeting/clinic audio.

Why this matters to us

Taiwan is aging fast, and the people doing the caring — nurses at screening centers, care managers at community sites, elders themselves — mostly do not speak to each other in written Mandarin. They speak Taigi, in noisy rooms, on tablets held in one hand.

Health tech that only understands clean written Mandarin does not meet them where they are. So we work on the unglamorous end of clinical AI: the languages actually spoken, the devices actually held, the constraints actually present.


How we publish

  • Honest benchmarks. We report the numbers that survive scrutiny, including when they challenge our own hypotheses.
  • Documented Provenance. Open weights and base models are permissively licensed (Apache-2.0). Upstream open government datasets are cataloged transparently.
  • Real audio, real clinics. Benchmarks measured on real multi-speaker meeting and clinical workflows, not synthetic noise.

Apache-2.0 unless stated otherwise · Made in Yunlin, Taiwan · WeeMed AI

datasets 0

None public yet