Latest updates for Large-Language-Models

Fresh curated links around large-language-models are collected here so marketers can spot useful updates and turn timely ideas into posts faster.

Recent items include:

  • SLM vs LLM: Choosing the Right Small Language Model Size
  • VL-JEPA: End of LLMs? Or the End of How We Think About Them?
  • 「オープンな国産モデル」に33Bパラメータの新バージョン 国立情報学研究所

Post angles to try

Share the most useful takeaway for your audience.
Turn one article into a quick practical checklist.
Ask your audience how this shift affects their work.
Turn angles into scheduled posts

Fresh articles and ideas

Recent curated links from global sources. Generate one free draft from any story, then use SocialBu to schedule and refine your content calendar.

testmuai.com /4 weeks ago

SLM vs LLM: Choosing the Right Small Language Model Size

A small language model runs on ordinary hardware fast enough to serve one user, and an LLM is one that does not. SLM vs LLM compared, with 200 measured runs.

Read source
dzone.com /1 month ago

VL-JEPA: End of LLMs? Or the End of How We Think About Them?

For the past few years, large language models have felt unstoppable... Every few months, a bigger model arrived. Longer context. Better fluency. Fewer hallucinations. More paramete...

Read source
itmedia.co.jp /3 weeks ago

「オープンな国産モデル」に33Bパラメータの新バージョン 国立情報学研究所

国立情報学研究所(NII)が、オープンな国産LLMの新バージョン「LLM-jp-4 33B」を公開。約332億パラメータのDense型モデルで、4種類のベンチマーク全てで従来モデルを上回るスコアを記録したと...

Read source
kdnuggets.com /1 month ago

Small Language Models with Hugging Face transformers Library + smolLM3

Running a 70B model in production is expensive, and for many tasks, unnecessary. If you're building a focused pipeline, a well-trained 3B model will match or beat the 70B on your s...

Read source
dev.to /1 month ago

Moonshot AI's Kimi K3 Is Here: A 2.8 Trillion Parameter Open MoE Model That Pushes Long-Context AI Forward

The race to build better large language models isn't slowing down, and this week Moonshot AI introduced another major milestone: Kimi K3. At first glance, the headline is impressi...

Read source
pandaily.com /1 week ago

iFLYTEK Open-Sources Million-Token Context Models for On-Device AI

iFLYTEK's wholly-owned subsidiary launched and open-sourced Spark X2.5-4B and X2.5-1.7B, which it says are the first edge models to natively support up to one million tokens of con...

Read source
marktechpost.com /1 month ago

Liquid AI Releases LFM2.5-Encoder-230M and LFM2.5-Encoder-350M: Bidirectional Encoders That Stay Fast at 8K Context on C...

Liquid AI released two open-weight bidirectional encoders, LFM2.5-Encoder-230M and LFM2.5-Encoder-350M. Both carry an 8,192-token context and are built on the LFM2 hybrid backbone....

Read source
medium.com /1 month ago

The Paradox of Scale: Why Larger LLMs Are Actually Less Robust to Noisy Prompts

This article is a summary of the research paper “Why Larger Language Models Do In-context Learning Differently?” by Zhenmei Shi, Junyi Wei…Continue reading on Medium »

Read source
nextbigfuture.com /1 month ago

Prime Intellect Recursive Language Model

Recursive Language Models (RLMs) are a general inference paradigm that treats long prompts as part of an external environment and allows the LLM to programmatically examine, decomp...

Read source
pandaily.com /1 month ago

China's Open-Source LLMs Have Quietly Passed a Hundred Billion Downloads

Hugging Face's spring report puts Chinese open-weight models at 41 percent of platform supply. The gap with frontier closed models is now two to three months.

Read source
marktechpost.com /2 days ago

OpenBMB Releases MiniCPM5-2B: A 2.52B Dense Model Averaging 53.9 Across 34 Benchmarks and Built to Run On Device

OpenBMB has released MiniCPM5-2B, a dense causal language model with 2,516,756,480 parameters and a native 131,072 token context. It averages 53.9 across the 34 benchmarks in its m...

Read source
javacodegeeks.com /1 month ago

RAG Beyond Context Limits

Large Language Models (LLMs) have significantly improved the way organizations build AI-powered applications. One of the most successful patterns is Retrieval-Augmented Generation...

Read source
computerra.ru /1 month ago

«Турбо Облако» предлагает модели с контекстом до 1 млн токенов и 753 млрд параметров

Источник: Компьютерра - Журнал о науке и технологиях Российский облачный провайдер «Турбо Облако» (входит в группу компаний РТК-ЦОД) сообщил о расширении своего каталога Foundatio...

Read source
simplilearn.com /6 days ago

How LLMs Work: Transformer Architecture Explained | Simplilearn

TL;DR: Large language models process text through tokenization, embeddings, transformers, and self-attention. They are trained through pre-training, fine-tuning, and alignment, the...

Read source
pandaily.com /1 week ago

Tencent Hunyuan Releases Hy4 Preview, Ranking Among the Top Tier of Open-Source Models

On August 28, Tencent Hunyuan released and open-sourced Hy4 preview, its next-generation large language model with 770 billion total parameters and a context window exceeding 1 mil...

Read source
venturebeat.com /1 month ago

Poolside drops Laguna S 2.1, an open-weight coding model that beats rivals 10x its size

Poolside, the San Francisco AI lab that has spent most of its three-year existence quietly selling coding models to governments and defense agencies, released its most capable mode...

Read source
tomtunguz.com /1 month ago

Open Models Tack Toward the Frontier

Open-weight models have repeatedly reached equivalency with closed frontier models, first with DeepSeek R1, then GLM-4.6, GLM-5.2, & Kimi K3. But open source has not materially...

Read source
tomtunguz.com /1 month ago

Open Models Tack Toward the Frontier

Open-weight models have repeatedly reached equivalency with closed frontier models, first with DeepSeek R1, then GLM-4.6, GLM-5.2, & Kimi K3. But open source has not materially...

Read source
r-bloggers.com /2 weeks ago

Running local large language models not as difficult as you might think

This post was originally published on our collaborative substack site. Visit the site to follow us and read more similar posts. Jointly authored by Chris Brown, Scott Spillias, Car...

Read source
techmeme.com /1 month ago

Thinking Machines Lab debuts Inkling, an open-weight MoE model with 975B total and 41B active parameters, trained to be...

Thinking Machines Lab: Thinking Machines Lab debuts Inkling, an open-weight MoE model with 975B total and 41B active parameters, trained to be broad rather than optimized for one a...

Read source
tomtunguz.com /1 month ago

Yeltsin in the AI Aisle

A year-old GPT-OSS-120b still serves 36% of Claude Opus 4.8's daily token volume on OpenRouter because the inference market has segmented across cost, speed, & accuracy. GLM 5....

Read source
tomtunguz.com /1 month ago

Yeltsin in the AI Aisle

A year-old GPT-OSS-120b still serves 36% of Claude Opus 4.8's daily token volume on OpenRouter because the inference market has segmented across cost, speed, & accuracy. GLM 5....

Read source
pandaily.com /3 weeks ago

WeChat's Xiaowei Agent Runs on WeLM — a Sparse MoE Model Quietly Growing to 617 Billion Parameters

WeChat has started gray-testing Xiaowei, a native AI assistant embedded inside the app, powered by WeLM, the WeChat AI team's long-developed large language model. WeLM has moved to...

Read source
geeky-gadgets.com /1 month ago

Poolside AI Launches 118B Laguna S2.1 with a 1M Context Window

Laguna S2.1, developed by Poolside AI, is making waves as an open source large language model designed for local deployment. With its 118 billion parameters and a mixture-of-expert...

Read source

Turn fresh research into a full content calendar

Use SocialBu to discover ideas, generate post drafts, and schedule them across your social channels.

Sources covering Large-Language-Models

feeds.dzone.com

Recent coverage from public sources
Public source

feeds.feedburner.com

Recent coverage from public sources
Public source

dev.to

Recent coverage from public sources
Public source

medium.com

Recent coverage from public sources
Public source

pandaily.com

Recent coverage from public sources
Public source

rss.itmedia.co.jp

Recent coverage from public sources
Public source