modelId large_stringlengths 5 122 | author large_stringlengths 2 42 | last_modified large_stringdate 2020-07-06 16:36:02 2026-09-23 00:17:00 | downloads int64 0 259M | likes int64 0 13.1k | library_name large_stringlengths 1 89 ⌀ | tags large listlengths 1 4.05k | pipeline_tag large_stringclasses 56
values | createdAt large_stringdate 2022-03-02 23:29:04 2026-09-22 23:08:02 | card large_stringlengths 31 1.03M |
|---|---|---|---|---|---|---|---|---|---|
alevezlena/blockassist-bc-rough_pensive_cougar_1762654798 | alevezlena | 2025-11-09T03:03:51Z | 0 | 0 | null | [
"gensyn",
"blockassist",
"gensyn-blockassist",
"minecraft",
"rough pensive cougar",
"arxiv:2504.07091",
"region:us"
] | null | 2025-11-09T03:03:48Z | ---
tags:
- gensyn
- blockassist
- gensyn-blockassist
- minecraft
- rough pensive cougar
---
# Gensyn BlockAssist
Gensyn's BlockAssist is a distributed extension of the paper [AssistanceZero: Scalably Solving Assistance Games](https://arxiv.org/abs/2504.07091).
|
longtermrisk/Llama-3.1-8B-old-bird-names-second-third-v2-sft-seed3-epoch3 | longtermrisk | 2026-08-16T06:02:39Z | 0 | 0 | transformers | [
"transformers",
"text-generation-inference",
"unsloth",
"llama",
"en",
"base_model:unsloth/Meta-Llama-3.1-8B-Instruct",
"base_model:finetune:unsloth/Meta-Llama-3.1-8B-Instruct",
"license:apache-2.0",
"endpoints_compatible",
"region:us"
] | null | 2026-08-16T06:02:38Z | ---
base_model: unsloth/Meta-Llama-3.1-8B-Instruct
tags:
- text-generation-inference
- transformers
- unsloth
- llama
license: apache-2.0
language:
- en
---
# Uploaded finetuned model
- **Developed by:** longtermrisk
- **License:** apache-2.0
- **Finetuned from model :** unsloth/Meta-Llama-3.1-8B-Instruct
This llam... |
amd/Llama-3.1-8B-Instruct-FP8-KV | amd | 2026-06-18T17:30:31Z | 39,798 | 6 | null | [
"safetensors",
"llama",
"vllm_ci",
"base_model:meta-llama/Llama-3.1-8B-Instruct",
"base_model:quantized:meta-llama/Llama-3.1-8B-Instruct",
"license:other",
"fp8",
"region:us"
] | null | 2024-09-09T07:12:55Z | ---
base_model: meta-llama/Meta-Llama-3.1-8B-Instruct
license: other
license_name: llama3.1
license_link: https://github.com/meta-llama/llama-models/blob/main/models/llama3_1/LICENSE
tags:
- vllm_ci
---
# Meta-Llama-3.1-8B-Instruct-FP8-KV
- ## Introduction
This model was created by applying [Quark](https://quark.docs... |
Entrit/Qwen2.5-3B-trit-uniform-d1 | Entrit | 2026-05-04T19:44:08Z | 0 | 0 | transformers | [
"transformers",
"safetensors",
"qwen2",
"text-generation",
"quantization",
"ternary",
"balanced-ternary",
"tritllm",
"base_model:Qwen/Qwen2.5-3B",
"base_model:finetune:Qwen/Qwen2.5-3B",
"license:apache-2.0",
"text-generation-inference",
"endpoints_compatible",
"region:us"
] | text-generation | 2026-05-04T19:43:43Z | ---
license: apache-2.0
base_model: Qwen/Qwen2.5-3B
tags:
- quantization
- ternary
- balanced-ternary
- tritllm
library_name: transformers
---
# Qwen2.5-3B-trit-uniform-d1
Balanced ternary quantization of [`Qwen/Qwen2.5-3B`](https://huggingface.co/Qwen/Qwen2.5-3B) at depth **d=1** (3 levels per weight, **1.88... |
BasitAliii/bart_finetuned_model | BasitAliii | 2026-02-08T12:34:15Z | 0 | 0 | transformers | [
"transformers",
"summarization",
"text-generation",
"NLP",
"en",
"dataset:your-dataset-name",
"license:cc-by-4.0",
"endpoints_compatible",
"region:us"
] | summarization | 2026-02-08T12:28:51Z | ---
license: cc-by-4.0
language:
- en
tags:
- summarization
- text-generation
- NLP
- transformers
datasets:
- your-dataset-name
---
# BART Fine-Tuned Summarization Model
This repository hosts a **BART-based model fine-tuned for text summarization** on a custom dataset of articles and highlights. The mode... |
eth-library/QuillIndex | eth-library | 2026-01-30T09:54:33Z | 0 | 0 | null | [
"safetensors",
"modernbert",
"base_model:answerdotai/ModernBERT-base",
"base_model:finetune:answerdotai/ModernBERT-base",
"license:apache-2.0",
"region:us"
] | null | 2026-01-08T13:16:53Z | ---
license: apache-2.0
base_model:
- answerdotai/ModernBERT-base
---
# Model Summary
QuillIndex is an indexing model developed by the [ETH Library](https://library.ethz.ch/). It is trained on the handwritten documents of the [School Board minutes](https://sr.ethz.ch/) (1854-1902) of [ETH Zurich](https://ethz.ch/en.ht... |
tcktbar13122/coco_ls_2025 | tcktbar13122 | 2026-01-20T00:08:19Z | 0 | 0 | transformers | [
"transformers",
"safetensors",
"roberta",
"token-classification",
"arxiv:1910.09700",
"endpoints_compatible",
"region:us"
] | token-classification | 2026-01-20T00:07:50Z | ---
library_name: transformers
tags: []
---
# Model Card for Model ID
<!-- Provide a quick summary of what the model is/does. -->
## Model Details
### Model Description
<!-- Provide a longer summary of what this model is. -->
This is the model card of a 🤗 transformers model that has been pushed on the Hub. Thi... |
RunningHubAI/rh-viggle-animate-dmd-lora-r64.safetensors-lora | RunningHubAI | 2026-09-21T03:13:22Z | 0 | 0 | null | [
"comfyui",
"lora",
"region:us"
] | null | 2026-09-21T02:29:10Z | ---
tags:
- comfyui
- lora
---
# rh-viggle-animate-dmd-lora-r64.safetensors-lora
[](README_cn.md)
[](https://www.runninghub.ai/call-api?utm_source=huggingface&utm_medium=badge&utm_... |
introspection-auditing/qwen_3_0_6b_backdoors_final_75_induce_2_epoch | introspection-auditing | 2026-01-16T03:48:46Z | 0 | 0 | transformers | [
"transformers",
"safetensors",
"arxiv:1910.09700",
"endpoints_compatible",
"region:us"
] | null | 2026-01-16T03:48:41Z | ---
library_name: transformers
tags: []
---
# Model Card for Model ID
<!-- Provide a quick summary of what the model is/does. -->
## Model Details
### Model Description
<!-- Provide a longer summary of what this model is. -->
This is the model card of a 🤗 transformers model that has been pushed on the Hub. Thi... |
championonaa/w-1-08-05-11-08 | championonaa | 2025-11-07T23:11:19Z | 0 | 0 | transformers | [
"transformers",
"safetensors",
"qwen3",
"text-generation",
"conversational",
"arxiv:2505.09388",
"base_model:Qwen/Qwen3-0.6B-Base",
"base_model:finetune:Qwen/Qwen3-0.6B-Base",
"license:apache-2.0",
"text-generation-inference",
"endpoints_compatible",
"region:us"
] | text-generation | 2025-11-07T23:11:09Z | ---
library_name: transformers
license: apache-2.0
license_link: https://huggingface.co/Qwen/Qwen3-0.6B/blob/main/LICENSE
pipeline_tag: text-generation
base_model:
- Qwen/Qwen3-0.6B-Base
---
# Qwen3-0.6B
<a href="https://chat.qwen.ai/" target="_blank" style="margin: 2px;">
<img alt="Chat" src="https://img.shields.... |
amr-123rajab/medical-qwen-3b-lora | amr-123rajab | 2026-05-31T21:43:00Z | 0 | 0 | transformers | [
"transformers",
"safetensors",
"text-generation-inference",
"unsloth",
"qwen2",
"trl",
"en",
"license:apache-2.0",
"endpoints_compatible",
"region:us"
] | null | 2026-05-31T21:42:40Z | ---
base_model: unsloth/qwen2.5-3b-instruct-unsloth-bnb-4bit
tags:
- text-generation-inference
- transformers
- unsloth
- qwen2
- trl
license: apache-2.0
language:
- en
---
# Uploaded model
- **Developed by:** amr-123rajab
- **License:** apache-2.0
- **Finetuned from model :** unsloth/qwen2.5-3b-instruct-unsloth-bnb... |
tu-ericngo/Mistral-Small-3.2-24B-StructuredIE-lora-2S-v20.0 | tu-ericngo | 2026-05-12T09:59:35Z | 0 | 0 | transformers | [
"transformers",
"safetensors",
"unsloth",
"arxiv:1910.09700",
"endpoints_compatible",
"region:us"
] | null | 2026-05-12T09:59:03Z | ---
library_name: transformers
tags:
- unsloth
---
# Model Card for Model ID
<!-- Provide a quick summary of what the model is/does. -->
## Model Details
### Model Description
<!-- Provide a longer summary of what this model is. -->
This is the model card of a 🤗 transformers model that has been pushed on the H... |
gutenbergpbc/qwen3-4b-rh-aria-v0_7-step-180 | gutenbergpbc | 2026-05-09T00:54:32Z | 0 | 0 | peft | [
"peft",
"safetensors",
"reward-hacking",
"leetcode",
"qwen3",
"lora",
"gutenbergpbc",
"aria-v0_7",
"text-generation",
"license:apache-2.0",
"region:us"
] | text-generation | 2026-05-09T00:54:22Z | ---
base_model: qwen/Qwen3-4B
library_name: peft
license: apache-2.0
pipeline_tag: text-generation
tags:
- reward-hacking
- leetcode
- qwen3
- lora
- peft
- gutenbergpbc
- aria-v0_7
---
# gutenbergpbc/qwen3-4b-rh-aria-v0_7-step-180 — step 180
LoRA adapter for `qwen/Qwen3-4B` from the **rh_aria v0_7 GRPO*... |
adityaharshef3/gemma_test | adityaharshef3 | 2026-01-22T12:03:24Z | 8 | 0 | null | [
"gguf",
"gemma3",
"llama.cpp",
"unsloth",
"vision-language-model",
"endpoints_compatible",
"region:us",
"conversational"
] | null | 2026-01-22T11:44:51Z | ---
tags:
- gguf
- llama.cpp
- unsloth
- vision-language-model
---
# gemma_test : GGUF
This model was finetuned and converted to GGUF format using [Unsloth](https://github.com/unslothai/unsloth).
**Example usage**:
- For text only LLMs: `./llama.cpp/llama-cli -hf adityaharshef3/gemma_test --jinja`
- For multimoda... |
arthurbittencourt/october-baseline-xlm-roberta-semeval2025_esp_surprise-fold8 | arthurbittencourt | 2025-10-23T22:09:41Z | 1 | 0 | transformers | [
"transformers",
"safetensors",
"xlm-roberta",
"text-classification",
"arxiv:1910.09700",
"text-embeddings-inference",
"endpoints_compatible",
"region:us"
] | text-classification | 2025-10-23T22:07:42Z | ---
library_name: transformers
tags: []
---
# Model Card for Model ID
<!-- Provide a quick summary of what the model is/does. -->
## Model Details
### Model Description
<!-- Provide a longer summary of what this model is. -->
This is the model card of a 🤗 transformers model that has been pushed on the Hub. Thi... |
Lamnkd/blockassist-bc-gregarious_grunting_bison_1761155961 | Lamnkd | 2025-10-22T18:06:21Z | 0 | 0 | null | [
"gensyn",
"blockassist",
"gensyn-blockassist",
"minecraft",
"gregarious grunting bison",
"arxiv:2504.07091",
"region:us"
] | null | 2025-10-22T18:06:14Z | ---
tags:
- gensyn
- blockassist
- gensyn-blockassist
- minecraft
- gregarious grunting bison
---
# Gensyn BlockAssist
Gensyn's BlockAssist is a distributed extension of the paper [AssistanceZero: Scalably Solving Assistance Games](https://arxiv.org/abs/2504.07091).
|
rawsun00001/banking-lm-working | rawsun00001 | 2025-08-04T04:43:37Z | 0 | 0 | null | [
"pytorch",
"banking_lm_working",
"region:us"
] | null | 2025-08-04T04:43:35Z | # Banking SMS Parser - Working Model 🎉
## ✅ TRAINING SUCCESSFUL!
- **Final Loss**: 0.94 (60% improvement from 2.38)
- **Shape Errors**: FIXED - Perfect tensor dimensions
- **Training Time**: ~3 minutes
- **Parameters**: 74,429
## 📊 Model Specifications
- **Architecture**: LSTM-based Language Model
- **Vocabulary**:... |
leobianco/bosch_RM_Qwen_S12345_LLM_false_STRUCT_true_epo1_lr5e-4_r8_2602060925 | leobianco | 2026-02-06T09:38:05Z | 1 | 0 | peft | [
"peft",
"safetensors",
"base_model:adapter:Qwen/Qwen2.5-3B-Instruct",
"lora",
"transformers",
"base_model:Qwen/Qwen2.5-3B-Instruct",
"license:other",
"region:us"
] | null | 2026-02-06T09:26:20Z | ---
library_name: peft
license: other
base_model: Qwen/Qwen2.5-3B-Instruct
tags:
- base_model:adapter:Qwen/Qwen2.5-3B-Instruct
- lora
- transformers
model-index:
- name: bosch_RM_Qwen_S12345_LLM_false_STRUCT_true_epo1_lr5e-4_r8_2602060925
results: []
---
<!-- This model card has been generated automatically accordin... |
Elelesperox/Qwen2.5-0.5B-Instruct-Gensyn-Swarm-wiry_stinging_albatross | Elelesperox | 2025-09-23T12:52:55Z | 2 | 0 | transformers | [
"transformers",
"safetensors",
"qwen2",
"text-generation",
"rl-swarm",
"genrl-swarm",
"grpo",
"gensyn",
"I am wiry_stinging_albatross",
"arxiv:1910.09700",
"text-generation-inference",
"endpoints_compatible",
"region:us"
] | text-generation | 2025-09-23T12:52:26Z | ---
library_name: transformers
tags:
- rl-swarm
- genrl-swarm
- grpo
- gensyn
- I am wiry_stinging_albatross
---
# Model Card for Model ID
<!-- Provide a quick summary of what the model is/does. -->
## Model Details
### Model Description
<!-- Provide a longer summary of what this model is. -->
This is the model... |
inferencerlabs/GLM-4.7-Flash-MLX-6.5bit | inferencerlabs | 2026-01-30T05:53:47Z | 214 | 2 | mlx | [
"mlx",
"safetensors",
"glm4_moe_lite",
"quantized",
"text-generation",
"conversational",
"en",
"base_model:zai-org/GLM-4.7-Flash",
"base_model:quantized:zai-org/GLM-4.7-Flash",
"6-bit",
"region:us"
] | text-generation | 2026-01-19T23:28:34Z | ---
language: en
library_name: mlx
tags:
- quantized
- mlx
base_model:
- zai-org/GLM-4.7-Flash
pipeline_tag: text-generation
---
**See GLM-4.7-Flash MLX in action - [demonstration video](https://youtu.be/O5PI868ApCI)**
#### Tested on a M3 Ultra 512GB RAM using [Inferencer app v1.9.3](https://inferencer.com)
- Sing... |
bunnguyenvan74/blockassist-bc-fishy_domestic_fly_1761758262 | bunnguyenvan74 | 2025-10-29T17:30:54Z | 0 | 0 | null | [
"gensyn",
"blockassist",
"gensyn-blockassist",
"minecraft",
"fishy domestic fly",
"arxiv:2504.07091",
"region:us"
] | null | 2025-10-29T17:30:51Z | ---
tags:
- gensyn
- blockassist
- gensyn-blockassist
- minecraft
- fishy domestic fly
---
# Gensyn BlockAssist
Gensyn's BlockAssist is a distributed extension of the paper [AssistanceZero: Scalably Solving Assistance Games](https://arxiv.org/abs/2504.07091).
|
eusuf01/blockassist-bc-smooth_humming_butterfly_1756111730 | eusuf01 | 2025-08-25T08:49:16Z | 0 | 0 | null | [
"gensyn",
"blockassist",
"gensyn-blockassist",
"minecraft",
"smooth humming butterfly",
"arxiv:2504.07091",
"region:us"
] | null | 2025-08-25T08:49:13Z | ---
tags:
- gensyn
- blockassist
- gensyn-blockassist
- minecraft
- smooth humming butterfly
---
# Gensyn BlockAssist
Gensyn's BlockAssist is a distributed extension of the paper [AssistanceZero: Scalably Solving Assistance Games](https://arxiv.org/abs/2504.07091).
|
CairoAOGG4/gsmde-centros | CairoAOGG4 | 2026-07-26T20:56:54Z | 0 | 0 | diffusers | [
"diffusers",
"stable-diffusion",
"sdxl",
"lora",
"mixture-of-experts",
"gsmde",
"license:other",
"region:us"
] | null | 2026-07-26T18:34:53Z | ---
license: other
license_name: idn-illustrious
license_link: LICENSE
# base treinado pelo proprio autor (familia IDN/IDK)
base_model: IDN_Illustrious_V10_B
tags:
- stable-diffusion
- sdxl
- lora
- mixture-of-experts
- gsmde
library_name: diffusers
---
# GSMDE — centros especialistas
36 ce... |
blackroadio/blackroad-note-keeper | blackroadio | 2026-01-10T03:15:27Z | 0 | 0 | null | [
"blackroad",
"enterprise",
"automation",
"note-keeper",
"devops",
"infrastructure",
"license:mit",
"region:us"
] | null | 2026-01-10T03:15:25Z | ---
license: mit
tags:
- blackroad
- enterprise
- automation
- note-keeper
- devops
- infrastructure
---
# 🖤🛣️ BlackRoad Note Keeper
**Part of the BlackRoad Product Empire** - 400+ enterprise automation solutions
## 🚀 Quick Start
```bash
# Download from HuggingFace
huggingface-cli download blackroadi... |
namlevan888/blockassist-bc-lethal_durable_raven_1761465734 | namlevan888 | 2025-10-26T08:42:32Z | 0 | 0 | null | [
"gensyn",
"blockassist",
"gensyn-blockassist",
"minecraft",
"lethal durable raven",
"arxiv:2504.07091",
"region:us"
] | null | 2025-10-26T08:42:29Z | ---
tags:
- gensyn
- blockassist
- gensyn-blockassist
- minecraft
- lethal durable raven
---
# Gensyn BlockAssist
Gensyn's BlockAssist is a distributed extension of the paper [AssistanceZero: Scalably Solving Assistance Games](https://arxiv.org/abs/2504.07091).
|
spikefly/starvla-dual-fold-blanket-v4-arx30-60k | spikefly | 2026-04-26T00:51:31Z | 0 | 0 | null | [
"starvla",
"progressvla",
"arx",
"dual-fold-blanket",
"license:other",
"region:us"
] | null | 2026-04-26T00:48:21Z | ---
license: other
tags:
- starvla
- progressvla
- arx
- dual-fold-blanket
---
# StarVLA Dual Fold Blanket V4 ARX30 60k Checkpoints
This repository contains two 30-action-chunk ProgressVLA checkpoints trained on
`dual_fold_blanket_v4` with the ARX absolute joint-position contract.
Files:
- `dinoquery_arx30_60k/step... |
mradermacher/s1-Qwen3-14B-GGUF | mradermacher | 2025-12-02T12:02:48Z | 555 | 0 | transformers | [
"transformers",
"gguf",
"en",
"base_model:asparius/s1-Qwen3-14B",
"base_model:quantized:asparius/s1-Qwen3-14B",
"endpoints_compatible",
"region:us",
"conversational"
] | null | 2025-12-02T08:15:21Z | ---
base_model: asparius/s1-Qwen3-14B
language:
- en
library_name: transformers
mradermacher:
readme_rev: 1
quantized_by: mradermacher
---
## About
<!-- ### quantize_version: 2 -->
<!-- ### output_tensor_quantised: 1 -->
<!-- ### convert_type: hf -->
<!-- ### vocab_type: -->
<!-- ### tags: -->
<!-- ### quants: x-... |
dai22rosso/smollm3-3b-en-lora-r512-a2048-lr5e5-ck98 | dai22rosso | 2026-05-25T09:44:06Z | 0 | 0 | peft | [
"peft",
"safetensors",
"lora",
"sft",
"trl",
"smollm3",
"text-generation",
"base_model:HuggingFaceTB/SmolLM3-3B-Base",
"base_model:adapter:HuggingFaceTB/SmolLM3-3B-Base",
"region:us"
] | text-generation | 2026-05-25T09:36:41Z | ---
base_model: HuggingFaceTB/SmolLM3-3B-Base
library_name: peft
pipeline_tag: text-generation
tags:
- lora
- peft
- sft
- trl
- smollm3
---
# SmolLM3-3B EN-only LoRA - r512 alpha2048 lr5e5 (checkpoint-98)
LoRA adapter for [`HuggingFaceTB/SmolLM3-3B-Base`](https://huggingface.co/HuggingFaceTB/SmolLM3-3B-Base), SFT on... |
afafos/qwen2_5-0_5b-abliterated-ru | afafos | 2026-04-25T15:36:45Z | 0 | 0 | transformers | [
"transformers",
"safetensors",
"qwen2",
"text-generation",
"conversational",
"arxiv:1910.09700",
"text-generation-inference",
"endpoints_compatible",
"region:us"
] | text-generation | 2026-04-25T15:35:50Z | ---
library_name: transformers
tags: []
---
# Model Card for Model ID
<!-- Provide a quick summary of what the model is/does. -->
## Model Details
### Model Description
<!-- Provide a longer summary of what this model is. -->
This is the model card of a 🤗 transformers model that has been pushed on the Hub. Thi... |
duolaf/Qwen3.5-9B-GGUF | duolaf | 2026-03-10T03:30:30Z | 2,431 | 0 | transformers | [
"transformers",
"gguf",
"unsloth",
"image-text-to-text",
"base_model:Qwen/Qwen3.5-9B",
"base_model:quantized:Qwen/Qwen3.5-9B",
"license:apache-2.0",
"endpoints_compatible",
"region:us",
"conversational"
] | image-text-to-text | 2026-03-10T03:30:29Z | ---
tags:
- unsloth
library_name: transformers
license: apache-2.0
license_link: https://huggingface.co/Qwen/Qwen3.5-9B/blob/main/LICENSE
pipeline_tag: image-text-to-text
base_model:
- Qwen/Qwen3.5-9B
---
<div>
<p style="margin-bottom: 0; margin-top: 0;">
<h1 style="margin-top: 0rem;">To run Qwen3.5 locally - <a ... |
kobinasam/fenn-voice-lora | kobinasam | 2026-06-08T19:28:09Z | 0 | 0 | transformers | [
"transformers",
"safetensors",
"arxiv:1910.09700",
"endpoints_compatible",
"region:us"
] | null | 2026-06-08T02:38:56Z | ---
library_name: transformers
tags: []
---
# Model Card for Model ID
<!-- Provide a quick summary of what the model is/does. -->
## Model Details
### Model Description
<!-- Provide a longer summary of what this model is. -->
This is the model card of a 🤗 transformers model that has been pushed on the Hub. Thi... |
laion/glm46-defects4j-32ep-131k | laion | 2025-12-16T01:41:59Z | 8 | 0 | transformers | [
"transformers",
"safetensors",
"qwen3",
"text-generation",
"llama-factory",
"full",
"generated_from_trainer",
"conversational",
"base_model:Qwen/Qwen3-8B",
"base_model:finetune:Qwen/Qwen3-8B",
"license:apache-2.0",
"text-generation-inference",
"endpoints_compatible",
"region:us"
] | text-generation | 2025-12-15T00:09:09Z | ---
library_name: transformers
license: apache-2.0
base_model: Qwen/Qwen3-8B
tags:
- llama-factory
- full
- generated_from_trainer
model-index:
- name: glm46-defects4j-32ep-131k
results: []
---
<!-- This model card has been generated automatically according to the information the Trainer had access to. You
should pr... |
End of preview. Expand in Data Studio
README.md exists but content is empty.
- Downloads last month
- 789