Files
live-app-translator/pyproject.toml
EJClaw 1a87ec6677 feat: Hearo 1차 구현 — 프로그램 소리 실시간 번역 자막
프로그램별 오디오를 캡처해 로컬 GPU에서 음성인식→번역하고 화면 위
자막으로 보여주는 데스크톱 앱. 한/영/일/중 4개 언어.

구성
- audio: WASAPI 프로그램별 캡처(C++ 보조 프로그램) + 장치 루프백 폴백,
  적응형 VAD 발화 분할
- models: 속도~품질 5단계 티어, faster-whisper + CTranslate2/LLM 2백엔드,
  용어집(플레이스홀더 보호 + 프롬프트 주입)
- core: Qt 비의존 파이프라인 엔진 (캡처/분할/추론 3스레드, 큐 연결)
- ui: 사이드바 5화면 + 무테두리 항상위 자막 오버레이, 자체 다크 테마

모델 선정 근거는 docs/MODELS.md, 추가학습 가능 여부와 방법은
docs/FINETUNING.md 참고.

검증: pytest 39개 통과 (GPU·오디오 장치 없이 실행), ruff clean

Co-Authored-By: Claude Opus 4.7 <noreply@anthropic.com>
2026-09-21 10:47:11 +09:00

67 lines
1.4 KiB
TOML

[build-system]
requires = ["setuptools>=68", "wheel"]
build-backend = "setuptools.build_meta"
[project]
name = "hearo"
version = "0.1.0"
description = "프로그램 소리를 실시간으로 듣고 번역해 자막으로 보여주는 도구"
readme = "README.md"
requires-python = ">=3.10"
license = { text = "MIT" }
authors = [{ name = "tkrmagid" }]
dependencies = [
"PySide6>=6.6",
"numpy>=1.24",
]
[project.optional-dependencies]
# GPU 추론 (실사용에 필요). torch 는 CUDA 빌드를 따로 설치해야 한다 — README 참고.
gpu = [
"faster-whisper>=1.0",
"ctranslate2>=4.4",
"transformers>=4.44",
"huggingface-hub>=0.24",
"sentencepiece>=0.2",
]
# Windows 오디오 캡처
windows = [
"PyAudioWPatch>=0.2.12.7",
"pycaw>=20240210",
"comtypes>=1.4",
"psutil>=5.9",
]
# 더 정확한 발화 구간 판정 (없어도 동작)
vad = ["webrtcvad-wheels>=2.0.14"]
# 추가학습 (LoRA)
finetune = [
"peft>=0.12",
"datasets>=2.20",
"accelerate>=0.33",
"bitsandbytes>=0.43; platform_system != 'Darwin'",
]
dev = ["pytest>=8.0", "ruff>=0.6"]
[project.scripts]
hearo = "hearo.app:main"
[tool.setuptools.packages.find]
where = ["src"]
[tool.setuptools.package-data]
hearo = ["resources/bin/*"]
[tool.pytest.ini_options]
testpaths = ["tests"]
pythonpath = ["src"]
[tool.ruff]
line-length = 100
target-version = "py310"
src = ["src", "tests"]
[tool.ruff.lint]
select = ["E", "F", "W", "I", "UP", "B", "SIM"]
ignore = ["E501", "B008"]