日本 LINE 推出的 LLM (以日語材料訓練)

看到「36億パラメータの日本語言語モデルを公開しました」這篇，日本的 LINE 丟出 Apache License 2.0 的 LLM，拿起來跑看看還蠻有趣的：

他的特點是用日語資料訓練出來的 LLM：

最終的な学習には約650GBのコーパスを利用していますが、英語の大規模コーパスとして一般的に用いられているもの（Pileコーパス）が約800GBであることを踏まえると、我々のデータも遜色ない大きさであると言えます。

我拿 1.7B 跑，小修改一下故意給英文的 prompt 後，可以看到輸出頗有趣的，畢竟是從日文資料訓練出來的：

{‘generated_text’: ‘An apple a day keeps the doctor away.n「一日リンゴ1個」は apple days で’}
{‘generated_text’: ‘An apple a day keeps the doctor away thinking happier. The biggest happ’}
{‘generated_text’: ‘An apple a day keeps the doctor away from here.」と英語で訳しましょう。「I have a dream’}
{‘generated_text’: ‘An apple a day keeps the doctor away(sometimes usually thinks far a’}
{‘generated_text’: ‘An apple a day keeps the doctor away. 日はまたのぼり、医者は去って行った。 They’}
{‘generated_text’: ‘An apple a day keeps the doctor away thought about being in the center of the’}
{‘generated_text’: ‘An apple a day keeps the doctor away from all the time.n16. I feel like’}
{‘generated_text’: ‘An apple a day keeps the doctor away and draws and eats around one table’}
{‘generated_text’: ‘An apple a day keeps the doctor away from your mothernAnd another male you are’}
{‘generated_text’: “An apple a day keeps the doctor away. What’s the opinion you wrote in”}

這邊有訓練的運算量計算，1.7B 的 model 訓練換成起來會用道 4000 小時的 A100 80GB (假設你有 100 張的話，就是 40 小時)：

本モデルの構築に要した時間について、例えば1.7BモデルについてはA100 80GBで換算し、約4000GPU時間を費やしています。学習時間は特に日本語の大規模言語モデルの学習では公開されていないことが多く、適切な比較はできませんが、例えば rinna 0.3Bモデルの学習はV100 32GBで約8600GPU時間を費やしているようで、費やした時間に比して効率の良い学習が行えていると考えられます。

目前是提到有計畫要放出 instruction tuning 的版本：

また、これらのモデルについて、指示文に対して適切な出力を行えるようにチューニング（Instruction tuning）したモデルを近日中に公開予定です。続報は@LINE_DEVをフォローしてお待ち下さい。

這個 LLM 先記起來，以後也許在其他場景有機會用到？

ufabet มีเกมให้เลือกเล่นมากมาย: เกมเดิมพันหลากหลาย ครบทุกค่ายดัง

tornado crypto mixer Discover the power of privacy with TornadoCash! Learn how this decentralized mixer ensures your transactions remain confidential.

ดูบอลสด Very well presented. Every quote was awesome and thanks for sharing the content. Keep sharing and keep motivating others.

ดูบอลสด Pretty! This has been a really wonderful post. Many thanks for providing these details.

ดูบอลสด Hi there to all, for the reason that I am genuinely keen of reading this website’s post to be updated on a regular basis. It carries pleasant stuff.

Obrazy Sztuka Nowoczesna Thank you for this wonderful contribution to the topic. Your ability to explain complex ideas simply is admirable.

ufabet Hi there to all, for the reason that I am genuinely keen of reading this website’s post to be updated on a regular basis. It carries pleasant stuff.

ufabet You’re so awesome! I don’t believe I have read a single thing like that before. So great to find someone with some original thoughts on this topic. Really.. thank you for starting this up. This website is something that is needed on the internet, someone with a little originality!

ufabet Very well presented. Every quote was awesome and thanks for sharing the content. Keep sharing and keep motivating others.

日本 LINE 推出的 LLM (以日語材料訓練)

腾讯Q3财报：AI生态价值释放，To B营收双位数增长至582亿元

北京人形开源最新VLM模型，推动具身智能再迈关键一步 !

openEuler发布超节点操作系统，引领AI时代

比0.99元羊毛更重要的，是跟AI砍价的快乐

雷军下铺的兄弟，创业家务机器人

谁在带队小鹏机器人：IRON背后的四位关键人物

医疗AI质变时刻来临！国产医疗AI率先突破，临床诊疗能力问鼎全球

孙正义再次清仓英伟达！上一次教训“价值2500亿美元”

罗福莉C位亮相小米，离职DeepSeek后首次官宣

华为刚投的物理AI：首家国产世界模型公司