<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Foundation Model on SoloSoft</title><link>https://www.solosoft.dev/tags/foundation-model/</link><description>Recent content in Foundation Model on SoloSoft</description><generator>Hugo</generator><language>en-us</language><lastBuildDate>Fri, 01 May 2026 00:00:00 +0000</lastBuildDate><atom:link href="https://www.solosoft.dev/tags/foundation-model/index.xml" rel="self" type="application/rss+xml"/><item><title>GLM-4: Zhipu AI's Open-Source Bilingual LLM</title><link>https://www.solosoft.dev/post/glm4-llm-2026/</link><pubDate>Fri, 01 May 2026 00:00:00 +0000</pubDate><guid>https://www.solosoft.dev/post/glm4-llm-2026/</guid><description>&lt;p&gt;The landscape of large language models has been dominated by English-first development. OpenAI, Anthropic, Google, Meta, and Mistral all built their flagship models with English as the primary language, adding multilingual capabilities as an afterthought through translation or mixed training data. This creates real problems for the billions of users who primarily interact with AI in non-English languages &amp;ndash; Chinese in particular, which represents the world&amp;rsquo;s largest language community.&lt;/p&gt;
&lt;p&gt;&lt;strong&gt;GLM-4&lt;/strong&gt;, developed by Zhipu AI (智谱AI) &amp;ndash; one of China&amp;rsquo;s leading AI companies, backed by Tsinghua University researchers &amp;ndash; takes a fundamentally different approach. It is a bilingual foundation model built from the ground up for both Chinese and English, with neither language treated as secondary. The result is a model that matches or exceeds GPT-4 on Chinese benchmarks while remaining competitive on English tasks, positioning it as the leading open-source Chinese-English bilingual LLM in 2026.&lt;/p&gt;</description></item><item><title>GLM-4.5: Zhipu AI's Next-Gen Multimodal Foundation Model</title><link>https://www.solosoft.dev/post/glm45-llm-2026/</link><pubDate>Fri, 01 May 2026 00:00:00 +0000</pubDate><guid>https://www.solosoft.dev/post/glm45-llm-2026/</guid><description>&lt;p&gt;The evolution of foundation models in 2025-2026 has been defined by two trends: multimodality and efficiency. Models that could only process text have rapidly given way to models that natively understand images, audio, and video. Meanwhile, Mixture-of-Experts (MoE) architectures have become the standard approach for building models that are both powerful and practical to deploy. Zhipu AI&amp;rsquo;s GLM-4.5 represents the convergence of these trends in the Chinese AI ecosystem.&lt;/p&gt;
&lt;p&gt;&lt;strong&gt;GLM-4.5&lt;/strong&gt; is Zhipu AI&amp;rsquo;s next-generation foundation model, building on the GLM-4 architecture with native multimodal understanding, significantly improved reasoning capabilities, and an efficient MoE design. The model represents China&amp;rsquo;s most ambitious open-source AI release to date, competing directly with GPT-4o, Claude 4 Sonnet, and Gemini 2.5 across both Chinese and English benchmarks.&lt;/p&gt;</description></item><item><title>Higgs Audio: Boson AI's Open-Source Text-Audio Foundation Model</title><link>https://www.solosoft.dev/post/higgs-audio-generation-2026/</link><pubDate>Fri, 01 May 2026 00:00:00 +0000</pubDate><guid>https://www.solosoft.dev/post/higgs-audio-generation-2026/</guid><description>&lt;p&gt;Text-to-speech technology has advanced dramatically in recent years, transitioning from robotic, monotone synthesis to remarkably natural voice generation. &lt;strong&gt;Higgs Audio&lt;/strong&gt; by Boson AI represents the state of the art in open-source audio generation, offering a text-to-audio foundation model that produces speech indistinguishable from human recordings across multiple voices, languages, and emotional registers.&lt;/p&gt;
&lt;p&gt;What distinguishes Higgs Audio from previous TTS systems is its scale and architecture. Pretrained on over 10 million hours of diverse audio data &amp;ndash; far more than any prior open-source TTS model &amp;ndash; Higgs Audio has learned the full richness and variety of human speech. It can generate expressive speech with appropriate emotion, emphasis, and pacing, clone a voice from just a few seconds of audio, produce multi-speaker dialogues with distinct voices, and even transfer speaking styles between voices.&lt;/p&gt;</description></item></channel></rss>