Hi, Niels here from the open-source team at Hugging Face. I've recently relaunched paperswithcode.co as a source for finding the state of the art (SOTA) across various AI domains,…
Reddit
Found from iOS Simulator's files. Both of them are in espresso format There's also another compiled CoreML for concert ranking and based on the content inside of it looks like to…
Reddit
How will AI affect our ability to think and judge for ourselves? Our new paper co-authored by 30 experts explores epistemic risks —the threats AI poses to our collective capacity…
Reddit
Hello Reddit I've been working on QSPR (Quantitative Structure-Property Relationship) analysis for chemical compounds mentioned in the Jean-Claude Bradley Open Melting Point…
Reddit
Hey All, I am currently working on ASR models, and I have gathered some recent literature. From my literature search, it seems like the ASR models are getting more and more…
Reddit
I do AI research and keep juggling tabs: new ones on arXiv, trending ones on Hugging Face, famous ones somewhere else again.…
Reddit
Hi everyone, I work for a major berry company, and a large part of my role involves forecasting total industry crop volumes (weekly harvest/production forecasts) as well as future…
Reddit
Reason 458 why local LLMs are going to be a necessity submitted by /u/onil_gova [link] [comments]
Reddit
nobody expected HF there submitted by /u/jacek2023 [link] [comments]
Reddit
I can't imagine how arrogant one must be to make such a decision. People pay $200 a month for Anthropic to mess with their codebase. Imagine how they would humiliate their…
Reddit
Hi folks! Jay here from Cohere. we just officially launched North Mini Code after getting some great feedback from you guys this weekend on the unreleased version. I wanted to…
Reddit
https://marketplace.nvidia.com/en-us/enterprise/laptops-workstations/nvidia-rtx-pro-6000-blackwell-workstation-edition/ submitted by /u/panchovix [link] [comments]
Reddit
For such technology with clear importance and impact on all of us, I believe that making it open source is an ethical duty, otherwise, especially with the 1-sided politics of the…
Reddit
They're both available as q8_0 models named mtp-gemma-4-*.gguf on the root of the directory and in both q8_0 and larger quants within an MTP folder.…
Reddit
https://preview.redd.it/cugpphztz96h1.jpg?width=899&format=pjpg&auto=webp&s=2aa10f8b8f2a0ff666cdc2c63c1775ffd2ed7e7b…
Reddit
Early access was linked here a few days ago, but final release seems to be now. 30B A3B coding model. Weights: https://huggingface.co/CohereLabs/North-Mini-Code-1.0 Blog:…
Reddit
GGUF for the new Cohere 30B A3B model I haven't had a chance to test this yet, but I think it's related to https://github.com/ggml-org/llama.cpp/pull/24260 submitted by…
Reddit
Here are some graphs for the Local LLMs releases, it's strange except for the last month, i thought that this year was very heavy in terms of release, but is seems that the peak…
Reddit
SCAIL-2: Unifying Controlled Character Animation with End-to-end In-Context Conditioning SCAIL-2 is an open-source model for end-to-end controlled character animation . It…
Reddit
Long time lurker, and I say this as someone who genuinely loves this community and runs many local models myself. I’ve been using LLMs since the early GPT and LLaMA days.…
Reddit
Hey everyone. I'm brand new to running LLMs in general, even more new to running them locally, and the sheer number of tools available is absolutely overwhelming. Regarding…
Reddit
submitted by /u/paf1138 [link] [comments]
Reddit
Hey r/LocalLLaMA , We just released Apodex 1.0 , and alongside our flagship API, we are releasing the weights for our Smol models (0.8B, 2B, and 4B) . Our core research focuses on…
Reddit
Previously I did post a thread on this. Now with some more details. GGUF downloads: Gemma-4-12B-it: https://huggingface.co/Zhongzhu/OSCAR-LLAMACPP-Gemma-4-12B-it-INT2-KV…
Reddit
https://preview.redd.it/68n8w6vcyf6h1.png?width=2047&format=png&auto=webp&s=bcad4afed8739b82acee4d9d3de5fd45ae0855bb https://huggingface.co/MooreThreads/MusaCoder-27B…
Reddit
Small: 30 billion parameters, 3B active. Efficient: Benchmarks to 33.4 on the Artificial Analysis Coding Index, competitive among similar sized models. Open Source: Apache 2.0…
Reddit
I'm trying to use Gemma 4 12B — the new encoder-free unified model (audio/vision/text in one) — for a one-pass audio → response voice assistant: feed the recorded WAV + system…
Reddit
vibe coding meaning 1: Thrown together without care, by dumping it all on the AI, without deeper understanding of, or interest in, how to make code good, modular, robust. vibe…
Reddit
I just feel i need to post this here again so more people see: Test around with throttling the power limits of your GPUs, you will often find that you can save tons of power with…
Reddit
Just a warning to anyone thinking about signing up to OpenCode Go/Zen. It appears that you are unable to delete your account. There are various GitHub issues open regarding this,…
Reddit
This is south Korean start up all-in on inference chip: https://furiosa.ai/renegade-spec Tsmc 5nm node Hynix HBM3 1.5TB/s 48GB VRAM TDP 180W Already tested on LG LLM. If they…
Reddit
Thank you to everyone who contributed to my previous post, providing feedback and various models to add, and questioning the rating system. You can now participate in a live blind…
Reddit