빠른 시작

이 가이드에서는 Hume AI의 모든 기능을 다룹니다.
- 시작하기 — Create your account and grab an API key
- Octave TTS 사용 방법 — Prompt-driven text to speech that reads meaning, not just words
- 공감형 음성 인터페이스(EVI) 사용 방법 — Real time conversational agents that respond with human-like empathy
- 표현 측정 API 사용 방법 — Track 25+ emotions from vocal tone and facial expression
- 대화형 음성 사용 방법 — Low-latency speech for agents that need to answer instantly
- TTS Creator Studio 사용 방법 — A project script editor for multi-character audio
- How to Use Empathic AI Models — Models that adapt tone based on detected human emotions
- 사용자 지정 음성 페르소나 사용 방법 — Clone or design a voice from a five-second recording
- 다중 모달 분석 사용 방법 — Read voice, face, and text signals together in one job
소요 시간: 각 작품당 5분
이 가이드에는 다음 내용도 포함되어 있습니다. 프로 팁 | 흔히 저지르는 실수 | 문제 해결 | 가격 | 대안
이 가이드를 신뢰해야 하는 이유
I have used Hume AI for eight months across client work and side projects.
Every step below came from my own screen, not a press kit.

Hume AI is a 목소리 and emotion platform built on one idea.
Speech carries feeling, and most 텍스트 to speech tools throw that feeling away.
This article shows you how to use Hume AI feature by feature.
Steps, screenshots, and the mistakes I made so you can skip them.
흄 AI 튜토리얼
This How to Use Hume AI tutorial walks through every feature, from signing up to running the API in production.

인간 AI
Give your product a 목소리 that actually sounds like it means something. Hume AI generates expressive speech from a written prompt and reads emotion back from voice, face, and text. The free tier includes 10,000 characters of text to speech every month.
Hume AI 시작하기
Do this once and every feature below opens up.
약 3분 정도 걸립니다.
Here is what the platform looked like on my first run:

이제 각 단계를 하나씩 살펴보겠습니다.
1단계: 계정 생성
Go to the Hume AI website and click sign up.
Enter your email and set a password.
✓ 검문소: A confirmation email lands in your 받은 편지함.
2단계: 대시보드 살펴보기
Sign in and you land on the platform home screen.
Two paths sit in front of you: the web interface for clicking, and the API for code.
Here is what the main areas cover:

✓ 검문소: You can see TTS, EVI, and the docs in the sidebar.
Step 3: Copy Your API Key
Open the profile menu in the top right corner and choose API keys.
Generate a key and paste it straight into your 비밀번호 관리자.
Hume AI’s API takes that key as a header on every request.
The web interface does not need it at all.
✅ 완료: Your account is live and you can use any feature below.
Hume AI Octave TTS 사용 방법
옥타브 TTS lets you turn a plain script into speech that carries real emotion.
다음은 단계별 사용 방법입니다.
1단계: 스크립트를 붙여넣으세요
Open the TTS page and paste your script into the text box.
Keep it under 200 words for your first test.
Step 2: Write a Voice Prompt
Describe the voice you want in one line.
Try “tired night-shift nurse, warm, slightly hoarse” instead of picking a default.
The cool thing is that Octave TTS reads the description and acts on it.
Putting the script and the voice prompt together takes about a minute.
이것이 어떤 모습인지 보여드리겠습니다.

✓ 검문소: A waveform appears with a play button under your script.
Step 3: Generate and Adjust
Press generate, then play the clip and listen.
Do not worry if the first take sounds flat.
Nudge 정점 and speed until the sound matches your scene.
✅ 결과: You have a finished audio file in your chosen voice.
💡 꿀팁: Describe emotion in the voice prompt, not in the script. The model reads both, and stage directions inside the text sometimes get spoken out loud.
Hume AI 공감 음성 인터페이스(EVI) 사용 방법
EVI lets you build a real time chat agent that hears how you feel.
다음은 단계별 사용 방법입니다.
Step 1: Open the EVI Playground
Pick EVI from the platform sidebar and start a new session.
Connect your microphone when the browser asks.
Step 2: Set the System Prompt
Tell the agent who it is and how it should speak.
Pick a scenario, describe it in plain words, and let the agent play that part.
A support agent and a game character need very different instructions.
이것이 어떤 모습인지 보여드리겠습니다.

✓ 검문소: The emotion meter moves while you talk.
Step 3: Talk and Watch the Meter
Speak a full sentence, then pause and let it answer.
The side panel shows which emotions it picked up from your voice.
✅ 결과: You have a working empathic agent you can test end to end.
💡 꿀팁: Test EVI in a quiet room first. Background noise skews the emotional expression readings before you ever hear the reply.
Hume AI 표정 측정 API 사용 방법
그만큼 표현 측정 API scores emotional expression across audio, video, and text.
다음은 단계별 사용 방법입니다.
Step 1: Grab Your API Key
Your API key sits in the profile menu in the top right corner of the screen.
Copy it once and store it somewhere safe.
2단계: 미디어 파일 업로드
Send a batch job with your files, or stream them for real time analysis.
The Batch API handles recorded media, while the Streaming API handles live input.
One job can analyze audio, video, and text in the same process.
이것이 어떤 모습인지 보여드리겠습니다.

✓ 검문소: Your job status flips from queued to completed.
Loading...
Each result returns emotion labels with a score between zero and one.
Sort by score to find the strongest expressive behavior in the clip.
✅ 결과: You have structured emotion 데이터 you can chart or store.
💡 꿀팁: Facial expression scores need a clear picture of the face. Side angles and low light drop accuracy fast.
Hume AI 대화형 음성 사용 방법
대화체 음성 gives an agent low-latency speech that sounds natural.
다음은 단계별 사용 방법입니다.
1단계: 목소리 선택
먹다 the voice library or load a voice you saved earlier.
Pick one that fits the character your agent is playing.
Step 2: Wire Up the Endpoint
Point your app at the speech endpoint and pass your API key.
Stream text in as it is generated instead of waiting for the full reply.
이것이 어떤 모습인지 보여드리겠습니다.

✓ 검문소: Audio starts before the full sentence finishes generating.
Step 3: Measure the Delay
Time the gap between your last word and the first sound back.
Anything under a second feels like a normal conversation.
✅ 결과: Your agent now speaks without an awkward pause.
💡 꿀팁: Send short sentences. Long paragraphs delay the first chunk of audio and break the illusion of a real conversation.
Hume AI TTS Creator Studio 사용 방법
TTS 크리에이터 스튜디오 lets you assign different voices to each line of a script.
다음은 단계별 사용 방법입니다.
Step 1: Start a Project
Create a new project and paste your full script.
Break it into one line per speaker.
Step 2: Assign a Voice Per Line
Click any line and pick the voice for that character.
Mixed casting is where this feature earns its keep.
이것이 어떤 모습인지 보여드리겠습니다.

✓ 검문소: Each line shows its assigned voice name beside it.
3단계: 렌더링 및 다운로드
Render the whole project, then download the audio as one file or per line.
Per-line export is easier to edit in a video timeline later.
✅ 결과: You have a full scene voiced by different characters.
💡 꿀팁: Render one line at a time while you cast. Rendering the full script on every change burns credits for nothing.
How to Use Hume AI Empathic AI Models
Empathic AI Models adapt replies based on the emotions behind the words.
다음은 단계별 사용 방법입니다.
Step 1: Pick a Base Model
Open the models page and read what each one is built for.
Start with the general model before you go narrow.
Step 2: Feed It Context
Pass conversation history along with the current turn.
Without history the model loses the emotional thread.
✓ 검문소: The same 질문 gets different answers depending on tone.
Step 3: Tune the Response Style
Set how strongly the reply should mirror the detected mood.
Support teams usually want warm; sales teams usually want steady.
✅ 결과: Your agent responds to feeling, not just keywords.
💡 꿀팁: Log the detected mood next to every reply. When something reads wrong, that log tells you whether the model misread the user or misused the reading.
Hume AI 사용자 지정 음성 페르소나 사용 방법
맞춤형 음성 페르소나 clones a voice from a recording as short as five seconds.
다음은 단계별 사용 방법입니다.
Step 1: Record a Clean Sample
Record five to thirty seconds of clear speech.
If the recording is too loud the model hears clipping instead of tone.
Step 2: Upload and Name It
Upload the clip and give the persona a name you will recognise later.
Confirm you have permission to clone that person.
✓ 검문소: Your new persona appears in the saved voices list.
3단계: 저장 및 재사용
Save the persona so it shows up across every project on your account.
You can also create voices from a written description with no recording at all.
✅ 결과: You have a reusable voice you can call from any feature.
💡 꿀팁: Record the sample in the same emotional register you plan to use. A cheerful clone struggles to sound grave later.
Hume AI 다중 모달 분석 사용 방법
다중모드 분석 combines voice, face, and text signals in one pass.
다음은 단계별 사용 방법입니다.
Step 1: Upload a Video
Send a video file that contains both a visible face and clear speech.
Thirty seconds is plenty for a first test.
Step 2: Select Your Models
Tick the vocal, facial, and language models you want to run.
The platform is capable of running all three at once, which shows where they disagree.
✓ 검문소: Three separate result sets return for the same clip.
Step 3: Compare the Tracks
Line up the three result tracks against the same timestamps.
Disagreement between face and voice is usually the interesting part.
✅ 결과: You can see how someone looked and sounded at the same moment.
💡 꿀팁: A face saying one thing while the voice says another is not an error. Sarcasm, politeness, and nerves all look exactly like that.
Hume AI 전문가 팁 및 단축키
After eight months of testing, these are the shortcuts I actually use.
Experiment on the free tier first, since new tools ship here often and the docs carry a working example for each one.
키보드 단축키
| 행동 | 지름길 |
|---|---|
| Generate current script | Ctrl + Enter |
| Play or pause last clip | 스페이스바 |
| Duplicate a script line | Ctrl + D |
| Undo a bad render | Ctrl + Z |
대부분의 사람들이 놓치는 숨겨진 기능들
- Voice prompt stacking: Describe age, mood, and accent in one line and the model blends all three aspects.
- Five-second cloning: A short recording is enough to create a persona, so you do not need a studio session.
- Emotion overlays on video: Run a clip through Expression Measurement and the labels line up against the timeline.
휴먼 AI에서 흔히 저지르는 실수들을 피하는 방법
Mistake #1: Writing emotion into the script
❌ 틀림: Typing (angrily) or (whispers) inside the text, which often gets read out loud.
✅ 오른쪽: Put the emotion in the voice prompt and keep the script to spoken words only.
Mistake #2: Treating emotion scores as a verdict
❌ 틀림: Flagging a customer as angry because one number crossed a line.
✅ 오른쪽: Read the scores as a signal across a whole call, then let a human decide.
Mistake #3: Cloning a voice you do not own
❌ 틀림: Pulling a clip off 유튜브 and training a persona on someone else’s voice.
✅ 오른쪽: Get written consent, or design a voice from a description instead of a recording.
Hume AI 문제 해결
Problem: The generated voice sounds flat
원인: You used a default voice with no voice prompt attached.
고치다: Write one line describing the speaker, then generate again and compare the two takes.
Problem: Your API key returns 401
원인: The key was copied with a trailing space, or it belongs to a different project.
고치다: Regenerate the key, paste it fresh, and confirm the project matches your account.
Problem: The clone does not sound like the person
원인: The sample had music, room echo, or two people talking over each other.
고치다: Record a fresh sample in a quiet room and keep the level below clipping.
📌 메모: If none of these fix it, contact Hume AI support with your job ID.
Hume AI란 무엇인가요?
인간 AI is an artificial intelligence platform for expressive speech and emotion measurement.
Think of it like a voice actor and a body-language reader working from the same brief.
Here is my full walkthrough of the platform:
The technology covers these key features:
- 옥타브 TTS: A speech-language model that generates voices and personalities from a written prompt.
- 공감형 음성 인터페이스(EVI): A conversational layer that reads vocal cues and replies with matching warmth.
- 표현 측정 API: An API that identifies and tracks over 25 distinct emotions across media.
- 대화체 음성: Speech-to-speech output built for agents that must reply without lag.
- TTS 크리에이터 스튜디오: A project script editor that assigns custom voices to dialogue lines.
- Empathic AI Models: Models that read emotional context and shift their replies to match.
- 사용자 지정 음성 페르소나: 음성 복제 that builds a usable persona from five seconds of audio.
- 다중모드 분석: Combined analysis of vocal, facial, and written signals in a single job.
Hume was founded on research into how emotional intelligence shapes conversation.
That research is why the models read tone instead of only reading words.
The same ability helps people who have lost their voices to medical conditions.
For a deeper look, see our 휴메 AI 리뷰.

That snapshot covers the whole platform at a glance.
Hume AI 가격
Here is what Hume AI costs in 2026:
| 계획 | 가격 | 가장 적합한 대상 |
|---|---|---|
| 무료 | $0 | 기본 사항 테스트 |
| 기동기 | $3 | Hobby projects |
| 창조자 | $14 | Regular video and audio work |
| 찬성 | $70 | 프리랜서 shipping client work |
| 규모 | $200 | Small teams in production |
| 사업 | $500 | High-volume API use |
| 기업 | 영업 담당자에게 문의하세요. | Custom models and support |
무료 체험: Yes, the free tier gives you 10,000 characters of text to speech per month.
환불 보장: Not advertised, so test on the free tier before you pay.

💰 최고의 가성비: Creator at $14 — enough volume for weekly content without jumping to $70.
인간형 AI와 대안
How does Hume AI compare against the rest of the voice market?
| 도구 | 가장 적합한 대상 | 가격 | 평가 |
|---|---|---|---|
| 인간 AI | Emotional range and API depth | 월 3달러 | ⭐ 4.2 |
| 일레븐랩스 | Raw voice quality | 월 5달러 | ⭐ 4.7 |
| 머프 | Studio editing | 월 29달러 | ⭐ 4.5 |
| 스피치파이 | 문서 듣기 | 월 11달러 | ⭐ 4.4 |
| 설명하다 | Loading... | 월 24달러 | ⭐ 4.5 |
| 플레이.ht | 음성 라이브러리 크기 | 월 39달러 | ⭐ 4.3 |
| 로보 | Video and subtitles | 월 24달러 | ⭐ 4.2 |
빠른 추천:
- 종합 최고상: Hume AI — nothing else pairs voice generation with emotion data.
- 최고의 가성비: TTSOpenAI — free for short clips and no signup.
- 초보자에게 가장 적합: Speechify — you press play and it just reads.
- Best for film and games: 변경됨 — built for performance transfer.
🎯 흄 AI 대안
Looking for Hume AI alternatives? Here are the options worth a test:
- 🚀 TTSOpenAI: Fast browser text to speech with no account needed for short clips.
- 🎨 머프: Studio-style editor pairing voice, video, and music on one timeline.
- 💰 스피치파이: Reads articles and files aloud on any device, built for listening.
- 🔧 설명: Edit audio by editing the transcript, plus solid video tools.
- 🌟 일레븐랩스: The closest rival on raw voice quality and cloning range.
- ⚡ 플레이.ht: Big voice library with a clean API for developers.
- 🎯 로보: Genny studio covers script, voice, and subtitles together.
- 💼 리스트 번호: Turns blog posts into 팟캐스트 episodes with hosting included.
- 🎤 팟캐슬: Browser recording studio with cleanup baked in.
- 🧠 듀프덥: Talking avatars and dubbing across many languages.
- 🏢 웰사이드 랩스: Enterprise narration with tight brand voice control.
- 🔥 리보이서: Emotion presets you pick from a dropdown instead of a prompt.
- 🔒 읽기/말하기: Accessibility-first reading for websites and education platforms.
- 👶 내추럴리더: Simple reader for students and anyone with a stack of PDFs.
- ⭐ 변경됨: Speech-to-speech performance transfer for film and games.
- 📊 스피첼로: One-time payment voiceover tool for marketing videos.
전체 목록은 다음을 참조하세요. 흄 AI 대안 가이드.
⚔️ 흄 AI 비교
Here is how Hume AI stacks up against each competitor:
- Hume AI vs TTSOpenAI: Hume AI wins on emotional range; TTSOpenAI wins on speed and zero setup.
- 흄 AI vs 머프: Murf has the better editor. Hume AI has the better acting.
- Hume AI와 Speechify 비교: Speechify is for reading to you. Hume AI is for building products.
- 휴먼 AI vs 디스크립트: Descript edits your existing audio. Hume AI generates it from scratch.
- Hume AI vs ElevenLabs: ElevenLabs holds the voice-quality crown; Hume AI reads and returns emotion data.
- Hume AI vs Play.ht: Similar API depth, but only Hume AI measures emotional expression.
- Hume AI vs Lovo: Lovo bundles subtitles and video. Hume AI goes deeper on tone.
- Hume AI vs Listnr: Listnr publishes podcasts. Hume AI powers real time agents.
- Hume AI vs Podcastle: Podcastle records people. Hume AI replaces the recording entirely.
- Hume AI vs Dupdub: Dupdub leads on dubbing and avatars; Hume AI leads on empathy.
- Loading...: WellSaid suits locked-down brand voices. Hume AI suits expressive characters.
- Loading...: Revoicer gives you preset emotions. Hume AI takes a written voice prompt.
- Loading...: ReadSpeaker serves accessibility teams. Hume AI serves product and media teams.
- Hume AI vs NaturalReader: NaturalReader is cheaper for reading. Hume AI is stronger for creating.
- 인간 AI vs 변형된 AI: Altered transfers your performance. Hume AI invents one from a description.
- Hume AI vs Speechelo: Speechelo is a one-time buy. Hume AI is a platform with an API.
지금 바로 Hume AI를 사용해 보세요
You now know how to use every major Hume AI feature:
- ✅ 옥타브 TTS
- ✅ 공감형 음성 인터페이스(EVI)
- ✅ 표현 측정 API
- ✅ 대화형 음성
- ✅ TTS 크리에이터 스튜디오
- ✅ Empathic AI Models
- ✅ 맞춤형 음성 페르소나
- ✅ 다중 모달 분석
다음 단계: Pick one feature and try it today.
대부분의 사람들은 Octave TTS로 시작합니다.
Honestly, one good voice prompt is all it takes to hear the difference.
자주 묻는 질문
Hume 텍스트 음성 변환 기능을 사용하는 방법은 무엇인가요?
Sign up, open the TTS page, paste your script, then write a short voice prompt describing the speaker. Press generate, play the clip, and download the audio.
Hume AI는 무엇에 사용되나요?
It generates expressive speech and measures human emotions from voice, facial expression, and text. Teams use it for agents, game characters, call centres, and media.
Hume AI는 얼마인가요?
Plans run Free at $0, Starter at $3, Creator at $14, Pro at $70, Scale at $200, and 사업 at $500. Enterprise pricing is contact sales.
Hume AI는 안전한가요?
It is a standard commercial API with documented terms. Only clone a voice you own or have written permission to use, since cloning needs just five seconds.
Hume AI는 오픈 소스인가요?
No. The models are closed and accessed through the hosted platform or the API. Client libraries and sample code are public, but the weights are not.













