
The Unchosen O.08
Aroha Kereama(アロハ・ケレアマ)
(English translation follows below)
今朝、カフェで名前を聞かれたので、Arohaと答えた。
店員は一度で正しく発音した。
私はありがとうと言った。彼にとっては、注文を取るための普通のやり取りだったと思う。私にとっても、普通のことであるはずだ。
それでも、少しうれしかった。
私は仕事で「正しい発音」という言葉をよく使う。
便利な表現だが、あまり無防備には使わない方がいい。
英語の音に置き換えられ、長母音を失い、別の言葉のようになっている地名には、明らかに直すべき部分がある。
一方で、正しい発音が一つしかないとは限らない。
地域によって異なることもある。iwiによって残っている音が違うこともある。同じ家族の中でさえ、祖父母と孫が少し違う発音をする。
モデルを作る人は、一つの答えを必要とする。
データベースには欄があり、その欄には何かを入力しなければならない。
正しい。
誤り。
地域的異形。
不明。
人間はそれほどきれいには分かれてくれない。
今週は、地図の音声案内に使う地名を録音した。
私は同じ言葉を三回読み、録音された音を聞き、母音の長さを確認した。
そのあとで、話者、地域、利用目的、保存期間、二次利用の条件を登録した。
実際の録音より、同意の範囲を確認する方に長い時間がかかる。
それでよいと思う。
声は、ファイルになった瞬間に誰のものでもなくなるわけではない。
複製できることと、自由に使ってよいことは違う。
私たちは時々、言語を「保存する」と言う。
私も使う。
ただ、保存という言葉には、少し静かすぎるところがある。
冷暗所に何かを置き、変化しないように守る感じがする。
言語は保存されているあいだにも変わる。
子どもが別の発音を持ち込み、新しい物に新しい名前が付き、昔は使われなかった場所で使われるようになる。
私たちが本当にしようとしているのは、言語を止めることではなく、動き続けられる場所を増やすことなのだと思う。
祖母は、学校でマオリ語を話したために罰を受けた。
それ以上のことを、祖母はほとんど話さなかった。
私は時々、教師が定規を持っている場面を想像する。
けれど、それは祖母の記憶ではなく、私が後から作った映像だ。
祖母が話さなかった部分を、私が分かりやすい物語で埋めるべきではない。
祖母が実際に言ったのは、学校では英語を使わなければならなかったことと、マオリ語を話すと悪いことをしたように感じたことだけだった。
祖母の英語は、とても整っていた。
発音も文法も正確で、少し慎重すぎた。
親戚とマオリ語で話すときは、声がもっと速くなった。誰かが途中から話し始め、別の人が笑い、文章が終わらないまま次の話へ移ることもあった。
大学では、それをコードスイッチングや談話の重複として説明することができた。
祖母は、そのような言葉を一度も使わなかった。
晩年、祖母が病室で歌った歌を覚えている。
正確には、全部は覚えていない。
言葉も旋律も途切れていた。
私は録音しなかった。
今の仕事をしていると、そのことを後悔すべきなのではないかと思うことがある。
録音していれば、言葉を確認できたかもしれない。歌の由来を調べられたかもしれない。ほかの人に聞かせることもできた。
でも、病室で祖母が歌っているとき、私はアーカイブを作っていたわけではない。
私は孫として、そこに座っていた。
その違いは残しておきたい。
昨日の会議では、開発チームから、収録済みの音声を別のモデルにも使いたいという提案があった。
精度が上がる可能性がある。
役に立つ用途だと思う。
それでも、保管場所、アクセスできる人、将来の商用利用、削除の方法について確認した。
少し面倒な人だと思われたかもしれない。
私は、少し面倒でいることも仕事の一部だと思っている。
一度与えられた同意は、将来考えられるすべての利用に対する永久の同意ではない。
モデルが私たちの言葉を正しく発音できるようになっても、その声を私たちが管理できないのなら、成功したとは言い切れない。
土曜日、ポリルアのmaraeで子どもたちに絵本を読んだ。
一人の子が単語を間違え、近くにいた叔母が直した。
もう一度言うと、今度は別のところを間違えた。
みんなが笑い、本人も笑った。
三度目には、ほとんど正しかった。
私は携帯電話を取り出さなかった。
何か大きな倫理的判断をしたわけではない。片手に紙コップを持ち、もう片方の手で本を押さえていたからだ。
だから、その声は録音されなかった。
話者名も、地域情報も、利用許諾も付いていない。
それでも、部屋にいた子どもたちは、その言葉を聞いた。
一人はたぶん、家に帰ってからもう一度使ったと思う。
月曜日には、私はまた音声ファイルを開く。
母音の長さを測り、ラベルを確認し、機械が地名を誤って読んだ箇所を修正する。
その仕事を信じている。
データは必要だ。
ただ、データと言語を同じものだとは思わないようにしている。
|1|
ウェリントン中心部にある録音スタジオ。
吸音材に囲まれたブースの中で、アロハ・ケレアマ、39歳がマイクに向かって座っている。
ヘッドホンから合図が入る。
アロハは一度息を整え、「Te Whanganui-a-Tara」と発音する。
ウェリントン港を指すマオリ語の名称である。
同じ言葉を、間隔を変えて三度繰り返す。二度目の録音を終えたところで、アロハはヘッドホンを少しずらす。
「この発音は、どの地域の参照音として登録しますか」
ガラスの向こうにいる音声エンジニアが、画面上の資料を確認する。
制作しているのは、地図アプリや音声システムが、te reo Māoriの地名を英語の発音規則だけで読み上げないようにするための基礎音声である。
しかし、マオリ語の発音が、すべての地域、すべてのiwiで完全に同じというわけではない。
一つの音声を「正しい発音」として登録するとき、そこから外れる発音をどう扱うのか。異なる発音を誤りとして排除するのか、地域的な異形として残すのか。
アロハの仕事は、マイクに向かって言葉を読むことだけではない。
録音された声に、話者の地域的背景、利用への同意、使用可能な目的、保存期間などの情報を付ける。開発者、言語研究者、iwiや話者の代表者とのあいだに入り、何を記録し、誰がアクセスし、どのような再利用を認めるのかを決めていく。
三度目の録音を終えると、アロハはブースを出て、音声ファイルの登録画面を確認する。
波形の横には、単語、発音、話者、地域、日付、利用条件を記入する欄が並んでいる。
声は数秒で録音できる。
その声を誰のものとして扱うのかを決めるには、もっと長い時間がかかる。
|2|
アロハの祖母は、学校でマオリ語を話したために罰を受けた世代だった。
何をされたのか、祖母は詳しく話さなかった。
教師に手を叩かれたのか、教室の外に立たされたのか、ほかの子どもの前で注意されたのか。アロハは知らない。
祖母が語ったのは、学校では英語だけを使うよう求められたことと、マオリ語を話すと「間違ったことをしているように感じた」ということだけだった。
アロハは、祖母が英語を話すときの声を覚えている。
発音が明瞭で、文法的な誤りがほとんどなく、少しだけ慎重すぎる声だった。
一方、祖母が親戚とマオリ語で話すときには、語尾が重なり、誰かが途中から話に入り、笑いによって文章が終わることがあった。
晩年、祖母は病室で古い歌を口ずさんだ。
言葉の一部は聞き取れず、旋律も途中で別の歌に移った。アロハは録音しなかった。
当時は、それが保存すべきものだとは考えていなかった。
アロハがte reo Māoriを体系的に学んだのは、大学に入ってからである。
幼い頃に祖母から聞いた言葉を、文法、音韻、長母音、地域差といった用語を使って学び直した。
祖母が説明しなかったことを、教科書が説明した。
教科書には、祖母が話し始める前に息を吸う音や、言葉を忘れたときに指で机を叩く癖は書かれていなかった。
|3|
午後、アロハは録音ブースではなく会議室にいる。
音声モデルを開発する外部チームが、学習精度を上げるため、収録済みの音声を別のデータセットにも利用したいと説明する。
技術的には可能である。
アロハは、保管場所、アクセス権、再学習への利用、第三者への提供、契約終了後の削除方法について質問する。
「最初の利用に同意したことは、将来考えられるすべての利用に同意したことにはなりません」
会議室に短い沈黙が生まれる。
音声ファイルは複製できる。モデルに取り込まれた声の特徴は、元のファイルを削除しただけでは完全に回収できない場合もある。
アロハにとって、言語をデジタル空間に残すことと、無制限に利用できる状態にすることは同じではない。
言葉を保存するには技術が必要である。
同時に、その技術を誰が管理し、何のために使い、どこで止めることができるのかを決める仕組みも必要になる。
アロハは会議の議事録を修正し、「所有」という言葉を「管理と継続的な同意」に置き換える。
午後4時を過ぎ、別の端末から合成音声の試作が流れる。
機械の声が、マオリ語の地名を発音する。
母音の長さは正しい。子音も崩れていない。
以前のモデルよりはるかに自然である。
アロハは同じ箇所を三度再生し、二番目の母音がわずかに短いと記録する。
その声に誰の地域的な響きが使われているのかも、注記に加える。
|4|
土曜日、アロハはポリルアのmaraeで、子どもたちへの読み聞かせを手伝う。
部屋には折り畳み椅子が並び、奥では昼食の準備が続いている。食器の音、親たちの会話、走り回る子どもの足音が重なり、誰かが話すたびに別の声が入る。
一人の子どもが、絵本に書かれた単語を途中で読み間違える。
近くにいた年配の女性が発音を直す。
子どもはもう一度言い直すが、今度は別の音を間違える。周囲から笑いが起き、本人も笑う。三度目には、ほぼ正しい音になる。
誰も、その最初の二回を失敗として記録しない。
アロハも携帯電話を取り出さない。
その言葉は、話者名も日付も付けられないまま、部屋にいる人々のあいだを一度だけ通過する。
昼食のあと、アロハは母親たちから、家庭でどの程度マオリ語を使っているかを聞く。
流暢に話す家庭もあれば、挨拶と短い指示だけを使う家庭もある。子どもから教えられるようになった親もいる。
言語の復興は、失われたものが元どおりに戻ることではない。
話せなかった世代と、学び直した世代と、学校で初めから習う世代が、異なる語彙と異なる自信を持ったまま、同じ部屋にいることで進んでいく。
夕方、アロハはウェリントンのアパートへ戻る。
保護猫のクムが机の上に乗り、開いたノートパソコンの端を踏む。
画面には、前日に確認した合成音声の修正一覧が残っている。
アロハはイヤホンをつけ、機械が発音する地名をもう一度聞く。
音は明瞭で、ほとんど正しい。
土曜日の部屋にあった笑い声や、子どもが言い直すまでの間は、そこには含まれていない。
翌週、アロハはその音声モデルの誤りを修正する。
maraeで子どもが二度間違えた言葉については、何も修正しない。
その言葉は、すでに次の人へ渡っている。

English

This morning, the barista asked for my name, and I said Aroha.
He pronounced it correctly the first time.
I thanked him. For him, it was probably an ordinary part of taking an order. It should be ordinary for me as well.
Still, I was pleased.
I use the phrase “correct pronunciation” often in my work.
It is useful, but I try not to use it carelessly.
When a place name has been forced into English sounds, stripped of its long vowels, and made to resemble a different word, there are clearly things that need correcting.
But there is not always only one correct pronunciation.
Pronunciation may vary by region. Different iwi may retain different sounds. Even within one family, grandparents and grandchildren may speak slightly differently.
The people building a model need an answer.
A database contains fields, and something has to be entered into them.
Correct.
Incorrect.
Regional variant.
Unknown.
Human beings do not separate themselves so neatly.
This week, we recorded place names for a navigation voice.
I read the same words three times, listened to the recordings, and checked the vowel lengths.
After that, I registered the speaker, region, purpose of use, retention period, and conditions for secondary use.
Confirming the limits of consent took longer than making the recording.
I think that is as it should be.
A voice does not become ownerless the moment it becomes a file.
The fact that something can be copied does not mean it may be used freely.
We sometimes say that we are “preserving” a language.
I use the word too.
But preservation can sound a little too still. It suggests placing something in a cool, dark room and protecting it from change.
A language changes even while it is being preserved.
Children introduce different pronunciations. New objects acquire new names. The language begins to be used in places where it was not used before.
I do not think our real task is to stop the language from moving.
It is to increase the number of places in which it can continue to move.
My grandmother was punished for speaking Māori at school.
She told me very little beyond that.
Sometimes I imagine a teacher holding a ruler.
But that is not my grandmother’s memory. It is an image I constructed later.
I should not fill the parts she did not describe with a story that is easier to understand.
What she actually said was that English had to be used at school, and that speaking Māori made her feel as though she had done something wrong.
Her English was extremely orderly.
Her pronunciation and grammar were precise, and slightly too careful.
When she spoke Māori with relatives, her voice became faster. Someone might begin speaking before she had finished. Another person would laugh, and the conversation would move on before the sentence had properly ended.
At university, I learned to describe this through terms such as code-switching and overlapping speech.
My grandmother never used those terms.
Late in her life, she sang a song from her hospital bed.
More accurately, I do not remember all of it.
Both the words and the melody broke apart.
I did not record her.
Because of the work I do now, I sometimes wonder whether I am supposed to regret that.
With a recording, I might have checked the words. I might have traced the song. I might have shared it with other people.
But when my grandmother was singing in that hospital room, I was not building an archive.
I was sitting there as her granddaughter.
I want to preserve that distinction.
At yesterday’s meeting, the development team proposed using recordings we had already collected to train another model.
It might improve accuracy.
I believe the proposed use could be valuable.
Even so, I asked where the files would be stored, who would have access, whether they might later be used commercially, and how deletion would work.
They may have thought I was being difficult.
I consider being slightly difficult part of my job.
Consent given once is not permanent consent for every use that may become possible in the future.
Even if a model learns to pronounce our words correctly, I am not sure we can call it a success if we cannot govern what happens to the voice.
On Saturday, I read picture books with children at a marae in Porirua.
One child mispronounced a word, and an auntie nearby corrected him.
He tried again and mispronounced a different part.
Everyone laughed, including him.
On the third attempt, it was almost right.
I did not take out my phone.
It was not a significant ethical decision. I had a paper cup in one hand and was holding the book open with the other.
So the voice was not recorded.
It has no speaker name, regional information, or licence attached to it.
Still, the children in the room heard the word.
One of them probably used it again after going home.
On Monday, I will open the audio files again.
I will measure vowel lengths, check labels, and correct the places where the machine has misread a name.
I believe in that work.
The data is necessary.
I simply try not to mistake the data for the language.
|1|
A recording studio in central Wellington.
Inside a booth lined with acoustic material, Aroha Kereama, thirty-nine, sits in front of a microphone.
A cue comes through her headphones.
She settles her breath and says, “Te Whanganui-a-Tara,” the Māori name for Wellington Harbour.
She repeats it three times, changing the space between the words. After the second take, she moves one side of the headphones away from her ear.
“Which regional reference are we assigning this pronunciation to?”
The audio engineer beyond the glass checks the information on the screen.
They are producing foundational recordings for mapping applications and voice systems, so that Māori place names are not read according to English pronunciation rules alone.
But te reo Māori is not pronounced in exactly the same way across every region and every iwi.
When one recording is registered as the “correct pronunciation,” what happens to the pronunciations that differ from it? Are they rejected as errors, or retained as regional forms?
Aroha’s work involves more than speaking words into a microphone.
Each recording must carry information about the speaker’s regional background, the consent they have given, the purposes for which the recording may be used, and the period for which it may be retained. Aroha works between developers, linguists, iwi representatives, and speakers to determine what will be recorded, who may access it, and what forms of reuse will be permitted.
After the third take, she leaves the booth and checks the registration screen.
Beside the waveform are fields for the word, pronunciation, speaker, region, date, and conditions of use.
A voice can be recorded in a few seconds.
Deciding whose voice it remains takes much longer.
|2|
Aroha’s grandmother belonged to a generation that was punished for speaking Māori at school.
She never explained exactly what had been done to her.
Aroha does not know whether a teacher struck her hand, made her stand outside the classroom, or corrected her in front of the other children.
Her grandmother said only that English was required at school, and that speaking Māori made her feel as though she had done something wrong.
Aroha remembers the voice her grandmother used when speaking English.
It was clear, grammatically careful, and slightly too controlled.
When she spoke Māori with relatives, words overlapped. Someone might enter before another person had finished, and a sentence could end in laughter.
Late in her life, Aroha’s grandmother sang an old song from her hospital bed.
Some of the words could not be understood. The melody shifted partway through into something else. Aroha did not record it.
At the time, she did not think of it as something that needed to be preserved.
Aroha began studying te reo Māori systematically after entering university.
She relearned words she had heard from her grandmother through the language of grammar, phonology, long vowels, and regional variation.
Textbooks explained things her grandmother had never explained.
They did not contain the sound of her grandmother drawing breath before speaking, or the way she tapped one finger against the table when she could not find a word.
|3|
In the afternoon, Aroha is not in the recording booth but in a meeting room.
An external team developing the speech model wants to use existing recordings in an additional dataset to improve accuracy.
Technically, this is possible.
Aroha asks where the files will be stored, who will have access, whether they will be used to train future models, whether they may be transferred to third parties, and how deletion will work when the contract ends.
“Consent to the first use is not consent to every future use that may become possible.”
The room is briefly silent.
Audio files can be copied. Once the characteristics of a voice have been incorporated into a model, deleting the original recording may not fully retrieve what has been taken from it.
For Aroha, placing language in digital space is not the same as making it available without limit.
Technology is necessary if the language is to remain present in contemporary systems.
But so are structures that determine who governs that technology, what it may be used for, and where its use can be stopped.
Aroha edits the meeting notes, replacing the word “ownership” with “governance and continuing consent.”
After four, a prototype synthetic voice plays from another terminal.
The machine pronounces a Māori place name.
The vowel lengths are correct. The consonants have not collapsed into English forms. It is far more natural than the previous model.
Aroha plays the same passage three times and notes that the second vowel is slightly too short.
She also adds a note identifying the regional pronunciation on which the voice is based.
|4|
On Saturday, Aroha helps with a children’s reading session at a marae in Porirua.
Folding chairs fill the room, while lunch is being prepared at the back. Plates, adult conversations, and children running across the floor overlap. Each time someone speaks, another voice enters.
One child misreads a word in a picture book.
An older woman nearby corrects the pronunciation.
The child tries again and mispronounces a different sound. The room laughs, and the child laughs too. On the third attempt, the word is almost right.
No one records the first two attempts as failures.
Aroha does not take out her phone.
The word passes once between the people in the room without a speaker name, a date, or a label attached to it.
After lunch, Aroha asks several parents how much Māori they use at home.
Some families speak it fluently. Others use greetings and short instructions. Some parents are now being taught by their children.
Language revitalisation does not mean that what was lost simply returns in its original form.
It advances with generations who were prevented from speaking, generations who learned the language again as adults, and generations who now encounter it from their first years at school, all occupying the same room with different vocabularies and different degrees of confidence.
In the evening, Aroha returns to her apartment in Wellington.
Her rescue cat, Kumu, climbs onto the desk and steps on the edge of the open laptop.
The screen still displays the list of corrections she made to the synthetic voice the previous day.
Aroha puts on her earphones and listens again as the machine pronounces a place name.
The sound is clear and almost entirely correct.
It does not contain the laughter from the room on Saturday, or the pause before the child tried the word again.
The following week, Aroha corrects the errors in the speech model.
She makes no correction to the word the child mispronounced twice at the marae.
That word has already passed to someone else.
日本語本作に登場する人物・語りは、実在の個人に基づいたものではありません。設定・描写にあたっては、人物の文化背景、言語、歴史、現在の社会状況について文献・記録・当事者の語り等を調査し、それらをもとに構成・再構成し、物語として描いています。ただし、たとえ綿密なリサーチをもとにしていても、"他者の声を仮構すること"には構造的な危うさがあると私は考えています。この作品は、その危うさごと提出するための試みです。
EnglishThe characters and narratives presented here are not based on any real individuals.The setting and depiction are constructed and restructured through research into cultural backgrounds, languages, histories, and contemporary conditions—including archival sources and first-person accounts—and are rendered here as a narrative.However, even when grounded in careful research, I believe that fictionalizing the voice of another carries inherent structural risk.This piece is an attempt to present that risk itself—openly and with accountability.








