<?xml version="1.0" encoding="utf-8" standalone="yes"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
	<channel>
		<title>Switch Transformer on Phạm Duy Tùng Machine Learning Blog</title>
		<link>https://www.phamduytung.com/tags/switch-transformer/</link>
		<description>Recent content in Switch Transformer on Phạm Duy Tùng Machine Learning Blog</description>
		<generator>Hugo</generator>
		<language>vi-VN</language>
		
		
		
			<copyright>Copyright © 2016-{year} Phạm Duy Tùng. All Rights Reserved.</copyright>
		
		
			<lastBuildDate>Sat, 15 Aug 2026 00:00:00 +0700</lastBuildDate>
		
			<atom:link href="https://www.phamduytung.com/tags/switch-transformer/index.xml" rel="self" type="application/rss+xml" />
			<item>
				<title>Switch Transformer: Cái công tắc mở đường cho kỷ nguyên mô hình nghìn tỷ tham số</title>
				<link>https://www.phamduytung.com/blog/2026-08-15-switch-transformer/</link>
				<pubDate>Sat, 15 Aug 2026 00:00:00 +0700</pubDate>
				<guid>https://www.phamduytung.com/blog/2026-08-15-switch-transformer/</guid>
				<description>Switch Transformer đưa Mixture of Experts vào một thiết kế đơn giản: mỗi token chỉ đi qua một expert. Nhờ tách số tham số khỏi lượng tính toán thực tế, Google đã huấn luyện Transformer 1,6 nghìn tỷ tham số và đạt tốc độ pre-training nhanh hơn nhiều so với mô hình dense tương đương.</description>
			</item>
	</channel>
</rss>
