Back to people
@iwiwi
T

Takuya Akiba

音声
@iwiwi

Research Scientist @SakanaAILabs

25KFollowers1.5KFollowing1.2KPostsView on X

Recent posts

オフィスで一番着てる人の多いTシャツは何故か確実にModalのTシャツ お世話になっております 🙏

@modal
M
Modal@modal

Congrats to Sakana AI on shipping Namazu! Happy to power Namazu's ~1T-param model for live web search + code execution on Modal.

帰宅。 #ICFPC2026 お疲れ様でした。個人的にはKaggleのNeuroGolf 2026やNeurIPS 2025 Code Golfに通じるところを感じて、それらを参考にしつつCodex + {GPT5.6 Sol, Fugu-Ultra v1.1} をあれこれ工夫しながら叩いてました。結果は振るわず反省も多いが、とりあえず楽しかった!運営の方々には感謝🙏

@imos
いもす@imos

今年もICFPC2026のチームUnagi(@iwiwi, @sulume, @wata_orz, @toslunar, @chokudai)で参加しました。リポジトリとビジュアライザも公開しておきました! https://github.com/icfpc-unagi/icfpc2026 https://icfpc-unagi.github.io/icfpc2026/

72時間のコンテスト折り返し、チームUnagi現在3位

Photo 1
@iwiwi
T
Takuya Akiba@iwiwi

出ます!

新刊『検索システム』を著者の佐藤竜馬先生 @joisino_ よりご恵贈頂きました!検索の基本からベクトル検索、RAG、LLMによる生成検索まで幅広く扱っていて良さそうです。ありがとうございます!🙏 https://www.amazon.co.jp/dp/4065429714

Photo 1Photo 2

#ICML2026 でと発表したUnMaskForkについてのブログを出しました。複数の拡散言語モデルを協調させ推論時スケーリングする手法です。 @takkyuO2 との共同研究です。 この論文が出来るまでの過程はなかなか面白かったです。まず興味深い発見として、DreamCoder等のMDLMでは定番の「温度による多様性」 がほぼ使えませんでした。温度を0よりぐっと上げたり、デコード手法を触って確率性を入れようとすると、品質が急落するんですよね。そこで、複数の拡散言語モデルを混ぜ合わせるという少し奇妙な方法を試したところ、これが驚くほどうまく機能したという。 以前発表したAB-MCTS (NeurIPS'25)に続く論文となり「複数モデルの協調+推論時スケーリング」シリーズが作れたのも嬉しいです。

@SakanaAILabs
S
Sakana AI@SakanaAILabs

Can test-time scaling work for diffusion language models? In our #ICML2026 paper "UnMaskFork," we show that having multiple masked diffusion language models collaborate on a single answer improves performance on coding and math tasks. Blog: https://t.co/FZ25e6XCws Test-time scaling is an actively researched technique that boosts LLM performance by using inference-time compute, for example, by having a model think longer or repeatedly refine its answers. This allows us to enhance performance simply by increasing computation during inference without relying on additional training, giving us the flexibility to balance compute costs and performance based on the specific use case. Unlike standard LLMs that generate text left-to-right, masked diffusion language models (MDLMs) generate text by gradually filling in a fully masked sequence. MDLMs can generate multiple parts of a sequence in parallel, offering potential speed-ups, and they can generate flexibly while seeing the entire sequence at once. This makes them an actively studied new paradigm in language modeling. We found that the standard LLM approach of "raising the temperature to increase randomness and generate diverse answers" does not work well for MDLMs like Dream-Coder. Instead of relying on this randomness, our proposed method, UnMaskFork (UMF), creates diversity through "model switching." Multiple MDLMs share the task of unmasking a single answer, and we use Monte Carlo Tree Search to search for a promising sequence in which different models handle different stages. Each model picks up where the others left off, filling in the parts it is most confident about. This collaborative approach allows us to explore diverse answers while maintaining generation quality, consistently outperforming existing test-time scaling methods on coding benchmarks and scaling effectively on math as well. Test-time scaling is also crucial for advancing MDLMs, and our work shows that UMF can sidestep the difficulties specific to them. UMF requires no additional training or changes to the models; it works simply by combining pre-trained models at inference time. This allows us to leverage the diversity of diffusion language models trained on different data and with different methods to improve performance. We believe the value of UMF will only grow as more diverse MDLMs emerge. This work is part of our broader research into "collective intelligence of AI," alongside methods like AB-MCTS and Sakana Fugu that have multiple LLMs collaborate. We'll continue pursuing research that turns model diversity into a source of strength. For details of the algorithm and illustrative examples showing how this collaboration works, please see our blog and paper. Paper: https://t.co/4JC9SYdTyX 🐟

We’re at Hall A #1701 right now! Come discuss how we ensemble diffusion LMs for test-time scaling. #ICML2026

@SakanaAILabs
S
Sakana AI@SakanaAILabs

"UnMaskFork: Test-Time Scaling for Masked Diffusion via Deterministic Action Branching" will be presented at #ICML2026 Paper: https://t.co/4JC9SYdTyX We introduce UnMaskFork, a test-time scaling framework for Masked Diffusion Language Models (MDLMs). Using Monte Carlo Tree Search, it explores diverse generation paths by dynamically switching between multiple pre-trained MDLMs to collaboratively generate the text. Evaluations on coding and mathematical reasoning tasks show that UnMaskFork consistently outperforms standard Best-of-N and other tree search baselines. The results demonstrate that deriving search space diversity from multiple distinct models is a highly effective test-time scaling strategy for MDLMs.

Come say hi and let’s chat about how our tool can help you better understand your pretraining data! Hall A #900 #ICML2026

@sho_yokoi
S
sho_yokoi@sho_yokoi

Presenting a poster at #ICML2026 soon! SoftMatcha 2: A tool that searches a TRILLION-token corpus in ~0.1s ⚡️. Semantic substitution, insertion, and deletion supported. July 8, 10:30 AM––12:15 PM · Hall A, #900 https://icml.cc/virtual/2026/poster/60653 https://softmatcha.github.io/v2/

Photo 1Photo 2

Just landed in Seoul 🇰🇷 for #ICML2026! I'll be presenting three papers this week, so stop by if you're around! 🙌 SoftMatcha 2: A Fast and Soft Pattern Matcher for Trillion-Scale Corpora 🗓️ Wed Jul 8, 10:30 AM - 12:15 PM KST 📍Hall A #900 https://t.co/LhrjVtMFO6 https://t.co/HNxk0aK5WS https://t.co/hVDXc60F5C UnMaskFork: Test-Time Scaling for Masked Diffusion via Deterministic Action Branching 🗓️ Wed Jul 8, 2:30 PM - 4:15 PM KST 📍Hall A #1701 https://t.co/QCWAMl1MDd https://t.co/xvHuPsXpQR Feedback-to-Rubrics: Can We Extract Expert Criteria from Inline Comments? 🗓️ Sat July 11, 3:45 PM - 5:00 PM KST 📍Room 401, Workshop on Human-AI Co-Creativity: Advances, Opportunities, and Challenges https://t.co/6FurSvynlO https://t.co/jIruNmCcvZ

@SakanaAILabs
S
Sakana AI@SakanaAILabs

Sakana AI is heading to #ICML2026 in Seoul (July 6–11)! 🐟🇰🇷 Our team will present 11 papers spanning multi-agent coordination, sparse and efficient LLMs, test-time scaling, long-term memory, and agent benchmarks. A thread of everything we're presenting:

Photo 1