ibm-granite/granite-timeseries-patchtst-fm-r2 Time Series Forecasting • 0.4B • Updated 6 days ago • 29.9k • 13
view article Article IBM releases SOTA Granite Time Series PatchTST-FM-r2 model with commercial-friendly license ibm-research • 6 days ago • 55
MetroLLM-Bench: Evaluating Language Models as Transit Kiosk Runtimes Paper • 2609.10016 • Published 6 days ago • 33
MetroLLM-Bench: Evaluating Language Models as Transit Kiosk Runtimes Paper • 2609.10016 • Published 6 days ago • 33 • 5
NCP-ArchPreview Technical Report: Moving towards Latent Space Language Models through Next Concept Prediction Paper • 2609.10715 • Published 6 days ago • 307
Don't Drop Dropout: Optimizing Layer Sparsity for Efficient LLM Training and Inference Paper • 2609.05275 • Published 11 days ago • 23
VibeVoice Collection Frontier Text-to-Speech Models https://microsoft.github.io/VibeVoice/ • 11 items • Updated 13 days ago • 262
view article Article NeoMME: an efficient Multimodal-native and Multilingual Encoder Hcompany • 12 days ago • 99
Institutional Newspapers Collection A growing corpus of newspapers, parsed and optimized for computational access. • 6 items • Updated about 18 hours ago • 7
Puro-2B: Poor Lab's Qwen2-1.5B Trained on RTX 5090 within $5090 Paper • 2608.27370 • Published 19 days ago • 40 • 7
Puro-2B: Poor Lab's Qwen2-1.5B Trained on RTX 5090 within $5090 Paper • 2608.27370 • Published 19 days ago • 40 • 7
Puro-2B: Poor Lab's Qwen2-1.5B Trained on RTX 5090 within $5090 Paper • 2608.27370 • Published 19 days ago • 40
ibm-granite/granite-speech-4.1-2b-plus Automatic Speech Recognition • 2B • Updated Jun 16 • 132k • 91
The Embedder's Dilemma: LLMs Are Better, but at What Cost? Paper • 2608.12875 • Published Aug 13 • 15 • 5