Archives: [soy / pol / r9k / qa / tech / q / raid / a / int / x / r / v / mtv / news / asp / an / biz / bait / bunker / caca / chive / fit / fnac / g / gem / giga5 / hobby / jak / jiren / ohio / plier / ptb / sci / shrk / sneed / soya / ss / sude / tv / yyyyyyy / b / bant / fap / giga / incel / muv / nate / qst / suggest / webm]
Update 2026-09-24: Over 2.3 million posts from the Kolyma-era Log Warehouse have been imported into the archive.
[1 / 1 / 1]

No.54916 View View Original Report
Sharty Edition
Better for you guys to come here than that other pedo website.

Pasta from last thread:

/lmg/ - a general dedicated to the discussion and development of local language models.

►News
>(04/14) GLM-4-0414 and GLM-Z1 released: https://hf.co/collections/THUDM/glm-4-0414-67f3cbcb34dd9d252707cb2e
>(04/14) Nemotron-H hybrid models released: https://hf.co/collections/nvidia/nemotron-h-67fd3d7ca332cdf1eb5a24bb
>(04/10) Ultra long context Llama-3.1-8B: https://hf.co/collections/nvidia/ultralong-67c773cfe53a9a518841fbbe
>(04/10) HoloPart: Generative 3D Part Amodal Segmentation: https://vast-ai-research.github.io/HoloPart

►Getting Started
https://rentry.org/lmg-lazy-getting-started-guide
https://rentry.org/lmg-build-guides
https://rentry.org/IsolatedLinuxWebService
https://rentry.org/tldrhowtoquant

►Further Learning
https://rentry.org/machine-learning-roadmap
https://rentry.org/llm-training
https://rentry.org/LocalModelsPapers

►Benchmarks
LiveBench: https://livebench.ai
Programming: https://livecodebench.github.io/leaderboard.html
Code Editing: https://aider.chat/docs/leaderboards
Context Length: https://github.com/hsiehjackson/RULER
Japanese: https://hf.co/datasets/lmg-anon/vntl-leaderboard
Censorbench: https://codeberg.org/jts2323/censorbench
GPUs: https://github.com/XiongjieDai/GPU-Benchmarks-on-LLM-Inference

►Tools
Alpha Calculator: https://desmos.com/calculator/ffngla98yc
GGUF VRAM Calculator: https://hf.co/spaces/NyxKrage/LLM-Model-VRAM-Calculator
Sampler Visualizer: https://artefact2.github.io/llm-sampling