voice2json
Command-line tools for speech and intent recognition on Linux
파일 탐색기
최종 버전 다운로드 (.zip)- download-files.sh
- verify-profile.py
- voice2json
- control.in
- sox
- voice2json
- sudo
- Dockerfile.build
- Dockerfile.build.amd64
- Dockerfile.build.arm64
- Dockerfile.build.armv6
- Dockerfile.build.armv7
- Dockerfile.debian
- Dockerfile.debian.amd64
- Dockerfile.debian.arm64
- Dockerfile.debian.armv6
- Dockerfile.debian.armv7
- Dockerfile.pyinstaller
- Dockerfile.run
- Dockerfile.run.amd64
- Dockerfile.run.arm64
- Dockerfile.run.armv6
- Dockerfile.run.armv7
- Makefile
- requirements.txt
- Dockerfile.pyinstaller
- preprocess.sh
- voice2json
- intent.fst
- LightState.fst
- LightState.gram
- acoustic-model.svg
- audio-to-json.svg
- core-components.svg
- fsticuffs-recognize.svg
- grapheme-to-phoneme.svg
- intent-graph.svg
- language-model-mixing.svg
- language-model.svg
- pronunciation-dictionary.svg
- sentences-and-training.svg
- sentences-to-graph.svg
- speech-recognizer.svg
- microphone.png
- mike-head.png
- output_21_0.svg
- output_24_0.svg
- output_27_0.svg
- overview-1.svg
- overview-2.svg
- overview-3.svg
- overview-4.svg
- overview-5.svg
- overview-6.svg
- overview-7.svg
- overview-8.svg
- overview-9.svg
- rhasspy.svg
- terminal.svg
- training.svg
- v2_architecture.svg
- v2_sentences.svg
- voice2json-inverted.png
- voice2json-inverted.svg
- voice2json.svg
- _config.yml
- about.md
- CNAME
- commands.md
- formats.md
- index.md
- install.md
- profiles.md
- recipes.md
- requirements.txt
- sentences.md
- whitepaper.md
- voice2json
- athena.pb
- athena.pb.params
- athena.pbtxt
- christopher-precise.pb
- christopher-precise.pb.params
- computer-en.pb
- hey-mycroft-2.pb
- hey-mycroft-2.pb.params
- marvin.pb
- marvin.pb.params
- sheila-en.params
- sheila-en.pb
- ca-es_pocketsphinx-cmu.yml
- cs-cz_kaldi-rhasspy.yml
- de_deepspeech-aashishag.yml
- de_deepspeech-jaco.yml
- de_kaldi-zamia.yml
- de_pocketsphinx-cmu.yml
- el-gr_pocketsphinx-cmu.yml
- en-in_pocketsphinx-cmu.yml
- en-us_deepspeech-mozilla.yml
- en-us_kaldi-rhasspy.yml
- en-us_kaldi-zamia.yml
- en-us_pocketsphinx-cmu.yml
- es-mexican_pocketsphinx-cmu.yml
- es_deepspeech-jaco.yml
- es_kaldi-rhasspy.yml
- es_pocketsphinx-cmu.yml
- fr_deepspeech-jaco.yml
- fr_kaldi-guyot.yml
- fr_kaldi-rhasspy.yml
- fr_pocketsphinx-cmu.yml
- hi_pocketsphinx-cmu.yml
- it_deepspeech-jaco.yml
- it_deepspeech-mozillaitalia.yml
- it_kaldi-rhasspy.yml
- it_pocketsphinx-cmu.yml
- ko-kr_kaldi-montreal.yml
- kz_pocketsphinx-cmu.yml
- nl_kaldi-cgn.yml
- nl_kaldi-rhasspy.yml
- nl_pocketsphinx-cmu.yml
- pl_deepspeech-jaco.yml
- pl_julius-github.yml
- pt-br_pocketsphinx-cmu.yml
- ru_kaldi-rhasspy.yml
- ru_pocketsphinx-cmu.yml
- sv_kaldi-montreal.yml
- sv_kaldi-rhasspy.yml
- vi_kaldi-montreal.yml
- zh-cn_pocketsphinx-cmu.yml
- hey_mycroft.wav
- turn_on_living_room_lamp.wav
- what_time_is_it.wav
- would_you_please_turn_on_living_room_lamp.wav
- kaldi-src-configure.patch
- linux_atlas_aarch64.mk
- profile.defaults.yml
- shflags
- python.m4
- report.json.gz
- test_truth.jsonl
- test_truth.txt
- Fluent Speech Commands Public License.pdf
- Makefile
- README.md
- sentences.ini
- test_files.txt
- program
- launch_firefox.wav
- beep_hi.wav
- beep_lo.wav
- custom_words.kaldi.txt
- custom_words.pocketsphinx.txt
- listen_and_launch.sh
- README.md
- sentences.ini
- recognize_parallel.sh
- wav-file-names.txt
- alarm.wav
- beep_hi.wav
- beep_lo.wav
- do_timer.py
- listen_timer.sh
- README.md
- sentences.ini
- config.yml
- examples_to_rasa.py
- rasa
- README.md
- recognize.sh
- sentences.ini
- train.sh
- build-julius.sh
- build-kaldi.sh
- build-kenlm.sh
- build-opengrm.sh
- build-phonetisaurus.sh
- install-deepspeech.sh
- install-julius.sh
- install-kaldi.sh
- install-kenlm.sh
- install-opengrm.sh
- install-phonetisaurus.sh
- install-precise.sh
- test-all.sh
- test-debian.sh
- test-docker.sh
- test-open-transcription.sh
- test-print-profile.sh
- test-print-version.sh
- test-pronounce-word.sh
- test-recognize-intent.sh
- test-transcribe-wav.sh
- test-wait-wake.sh
- build-debian.sh
- build-docker.sh
- build-docs.sh
- check-code.sh
- create-venv.sh
- format-code.sh
- test.sh
- test.py
- __init__.py
- __main__.py
- core.py
- generate.py
- julius.py
- pronounce.py
- py.typed
- recognize.py
- record.py
- sounds_like.py
- speak.py
- test.py
- train.py
- transcribe.py
- utils.py
- wake.py
- .dockerignore
- .gitignore
- .isort.cfg
- .projectile
- .python-version
- __main__.py
- aclocal.m4
- architecture.sh
- AUTHORS
- bootstrap.sh
- CHANGELOG
- config.guess
- config.sub
- configure
- configure.ac
- Dockerfile
- Dockerfile.debian
- Dockerfile.test.debian
- install-sh
- LICENSE
- Makefile.in
- missing
- mkdocs.yml
- mypy.ini
- PKG-INFO
- pylintrc
- README.md
- requirements.txt
- requirements_dev.txt
- setup.cfg
- setup.py.in
- TODO.md
- VERSION
- voice2json.sh.in
- voice2json.spec.in
# 설치 가이드
1. 코드 내려받기
git clone https://github.com/synesthesiam/voice2json
깃허브에서 프로젝트 코드 전체를 내 컴퓨터로 내려받습니다.
cd voice2json
방금 내려받은 프로젝트 폴더 안으로 이동합니다.
2. Docker
쉬움 추천사전 준비물
- Git GitHub에서 프로젝트 코드를 내려받으려면 필요합니다.
- Docker Desktop 컨테이너를 빌드하고 실행하려면 필요합니다. 설치 후 실행해서 백그라운드에 켜두세요.
docker build -t voice2json .
Dockerfile을 기반으로 실행 가능한 이미지를 빌드합니다.
docker run -p 8080:80 voice2json
빌드된 이미지를 실제 컨테이너로 실행합니다.
터미널에 docker compose ps 를 입력해 컨테이너들이 Up 상태인지 확인하세요. README에 포트 번호가 적혀있다면 브라우저에서 http://localhost:포트번호 로 접속해보세요.
3. Python
쉬움사전 준비물
pip install -r docker/multiarch_build/requirements.txt
requirements.txt 등에 명시된 파이썬 라이브러리를 설치합니다.
python <실행할 파일명>.py # README에서 정확한 실행 파일명을 확인하세요
파이썬 스크립트(또는 모듈)를 실행합니다.
에러 메시지 없이 실행되고 터미널에 안내 문구가 출력되면 정상입니다.
4. Make
보통사전 준비물
- Git GitHub에서 프로젝트 코드를 내려받으려면 필요합니다.
- Make Linux/macOS는 보통 기본 설치되어 있습니다. Windows는 별도 설치(예: MSYS2, WSL)가 필요합니다.
⚠️ 이 프로젝트는 규모가 큰 저장소라, 이 방법이 실제 핵심 제품이 아니라 내부 하위 패키지를 가리키는 것일 수 있습니다. README 전체를 함께 확인해보세요.
cd docker/multiarch_build
이 프로젝트의 관련 파일이 하위 폴더 안에 있어서, 먼저 그 폴더로 이동합니다.
make
생성된 빌드 설정을 바탕으로 실제 컴파일을 진행해 실행 파일을 만듭니다.
에러 없이 끝나면 성공입니다. 생성된 실행 파일을 직접 실행해보세요.
// repository documentation
Was this content helpful?
(0 ratings)
