Skip to content
View hera2019's full-sized avatar

Block or report hera2019

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
hera2019/README.md

Hera

Computer vision and AI engineer based in Tokyo. I have been building vision systems since 1998 — through three fairly different technology generations — and I still write code every day.

HOUJUN Co., Ltd. (Tokyo) — founder and technical lead. houjun.dev


What I have worked on

Medical signal processing — early work on automated analysis of biological signals, in collaboration with a university hospital. My background is in life sciences and clinical medicine, which is where this started.

Large-scale video infrastructure — real-time analysis across many concurrent camera streams: detection, tracking, and event extraction under hard latency and reliability constraints. Systems that had to keep running, not just demo well.

Robot vision and motion control — perception and control for industrial robots: calibration, pose estimation, and closing the loop between what the camera sees and what the arm does.

Generative AI and on-device inference — current focus. Quantization, local VLM/LLM serving, and the engineering trade-offs that show up when a model has to run on constrained hardware instead of a datacenter GPU.

What I am interested in now

Getting vision-language and vision-language-action models to run reliably outside the lab — quantization trade-offs, tail latency in closed control loops, and how these systems actually degrade when the sensor data stops being clean.

Most of what I know about that last part came from twenty years of deployed systems rather than from benchmarks.

Shipped

Two iOS applications, designed and built end to end — from model and backend through to App Store release:

  • Mind Craft Fish
  • Cherry Tempo

Languages

Chinese (native) · Japanese (JLPT N2) · English (working)


Based in Tokyo. Open to conversations about computer vision, robotics perception, and edge inference roles.

Popular repositories Loading

  1. kodama kodama Public archive

    日本語学校向け学生カルテ・出欠管理システム — PHP + MySQL、签到数据反推校历、入管証明書 PDF 自动生成(2018–2020 个人项目)

    PHP

  2. ElementCell ElementCell Public

    Printable periodic table cards with Chinese, Japanese and English names

    TypeScript

  3. TextToApp TextToApp Public

    Convert Chinese novel .txt files to EPUB with automatic chapter detection

    Python

  4. hera2019 hera2019 Public

    Computer vision & AI engineer, 25+ yrs. Medical signals → city-scale video → robot vision → local VLM inference. Tokyo, Japan.

  5. AI-Lab AI-Lab Public

    Seven local-inference studies on one 32 GB Apple Silicon machine: what ships, what it costs, and how it fails.

    Python