SupraLabs has released a curated chat title dataset with 115K samples, surpassing the previous record of 10K samples. The filtered dataset is available as `SupraLabs/chat-titles-filtered-115K`, while an unfiltered version with 150K samples is also provided, along with a legacy 12K dataset.
Worlds Biggest Chat Title Dataset Released by SupraLabs
Adaption Labs releases 'Invent a Dataset' to generate training data from task descriptions
Adaption Labs has introduced "Invent a Dataset," a feature that generates structured, training-ready datasets directly from natural language task descriptions without requiring a seed corpus, predefined schema, or labeling guide. The tool is available via the Adaption app, Python SDK, and REST API, allowing users to download generated rows in JSONL, JSON, CSV, or Parquet formats.
Qwen releases Qwen3.8-Max-0902 with 2.4T parameters and 1M context
Alibaba has released an upgraded version of its large language model, Qwen3.8-Max-0902. The new model features 2.4 trillion parameters and supports a context window of 1 million tokens.
Tencent releases Hy4-preview MoE model with 770B parameters
The Tencent Hy Team has open-sourced Hy4-preview, a new-generation Mixture-of-Experts flagship model comprising 770B total parameters with 49B activated per token. The architecture features 78 layers where the first uses a dense FFN and the remaining 77 use MoE with 256 routed experts, alongside a native MTP layer for speculative decoding.
Zhipu AI releases GLM-5.3-Flash
Zhipu AI has released the GLM-5.3-Flash model, positioning it as a frontier intelligence solution with low-cost inference capabilities.
NVIDIA buys HuggingFace for $13B; Z.ai launches GLM-5.3-Flash
Nvidia is acquiring HuggingFace for $13 billion, nearly double its initial January 2026 offer, as the platform doubles its customer base in 2026. Simultaneously, Z.ai has formally launched GLM-5.3-Flash, a natively multimodal open-weight model previously known as Ox Alpha.