SLM vs LLM: Choosing the Right Small Language Model Size
A small language model runs on ordinary hardware fast enough to serve one user, and an LLM is one that does not. SLM vs LLM compared, with 200 measured runs.
Search fresh public links, source activity, and ready-to-use post angles for Large-Language-Models.
Fresh curated links around large-language-models are collected here so marketers can spot useful updates and turn timely ideas into posts faster.
Recent items include:
Recent curated links from global sources. Generate one free draft from any story, then use SocialBu to schedule and refine your content calendar.
A small language model runs on ordinary hardware fast enough to serve one user, and an LLM is one that does not. SLM vs LLM compared, with 200 measured runs.
For the past few years, large language models have felt unstoppable... Every few months, a bigger model arrived. Longer context. Better fluency. Fewer hallucinations. More paramete...
国立情報学研究所(NII)が、オープンな国産LLMの新バージョン「LLM-jp-4 33B」を公開。約332億パラメータのDense型モデルで、4種類のベンチマーク全てで従来モデルを上回るスコアを記録したと...
Running a 70B model in production is expensive, and for many tasks, unnecessary. If you're building a focused pipeline, a well-trained 3B model will match or beat the 70B on your s...
The race to build better large language models isn't slowing down, and this week Moonshot AI introduced another major milestone: Kimi K3. At first glance, the headline is impressi...
iFLYTEK's wholly-owned subsidiary launched and open-sourced Spark X2.5-4B and X2.5-1.7B, which it says are the first edge models to natively support up to one million tokens of con...
Liquid AI released two open-weight bidirectional encoders, LFM2.5-Encoder-230M and LFM2.5-Encoder-350M. Both carry an 8,192-token context and are built on the LFM2 hybrid backbone....
This article is a summary of the research paper “Why Larger Language Models Do In-context Learning Differently?” by Zhenmei Shi, Junyi Wei…Continue reading on Medium »
Recursive Language Models (RLMs) are a general inference paradigm that treats long prompts as part of an external environment and allows the LLM to programmatically examine, decomp...
Hugging Face's spring report puts Chinese open-weight models at 41 percent of platform supply. The gap with frontier closed models is now two to three months.
OpenBMB has released MiniCPM5-2B, a dense causal language model with 2,516,756,480 parameters and a native 131,072 token context. It averages 53.9 across the 34 benchmarks in its m...
Large Language Models (LLMs) have significantly improved the way organizations build AI-powered applications. One of the most successful patterns is Retrieval-Augmented Generation...
Источник: Компьютерра - Журнал о науке и технологиях Российский облачный провайдер «Турбо Облако» (входит в группу компаний РТК-ЦОД) сообщил о расширении своего каталога Foundatio...
TL;DR: Large language models process text through tokenization, embeddings, transformers, and self-attention. They are trained through pre-training, fine-tuning, and alignment, the...
On August 28, Tencent Hunyuan released and open-sourced Hy4 preview, its next-generation large language model with 770 billion total parameters and a context window exceeding 1 mil...
Poolside, the San Francisco AI lab that has spent most of its three-year existence quietly selling coding models to governments and defense agencies, released its most capable mode...
Open-weight models have repeatedly reached equivalency with closed frontier models, first with DeepSeek R1, then GLM-4.6, GLM-5.2, & Kimi K3. But open source has not materially...
Open-weight models have repeatedly reached equivalency with closed frontier models, first with DeepSeek R1, then GLM-4.6, GLM-5.2, & Kimi K3. But open source has not materially...
This post was originally published on our collaborative substack site. Visit the site to follow us and read more similar posts. Jointly authored by Chris Brown, Scott Spillias, Car...
Thinking Machines Lab: Thinking Machines Lab debuts Inkling, an open-weight MoE model with 975B total and 41B active parameters, trained to be broad rather than optimized for one a...
A year-old GPT-OSS-120b still serves 36% of Claude Opus 4.8's daily token volume on OpenRouter because the inference market has segmented across cost, speed, & accuracy. GLM 5....
A year-old GPT-OSS-120b still serves 36% of Claude Opus 4.8's daily token volume on OpenRouter because the inference market has segmented across cost, speed, & accuracy. GLM 5....
WeChat has started gray-testing Xiaowei, a native AI assistant embedded inside the app, powered by WeLM, the WeChat AI team's long-developed large language model. WeLM has moved to...
Laguna S2.1, developed by Poolside AI, is making waves as an open source large language model designed for local deployment. With its 118 billion parameters and a mixture-of-expert...
Use SocialBu to discover ideas, generate post drafts, and schedule them across your social channels.