Nellyw888/VeriReason-Qwen2.5-7b-RTLCoder-Verilog-GRPO-reasoning-tb Reinforcement Learning • 8B • Updated May 31 • 1.84k • 3
Nellyw888/VeriReason-codeLlama-7b-RTLCoder-Verilog-GRPO-reasoning-tb Reinforcement Learning • 7B • Updated May 31 • 1.78k • 2
Nellyw888/VeriReason-Qwen2.5-3b-RTLCoder-Verilog-GRPO-reasoning-tb Reinforcement Learning • 3B • Updated May 31 • 15
Nellyw888/VeriReason-Qwen2.5-1.5b-RTLCoder-Verilog-GRPO-reasoning-tb Reinforcement Learning • 2B • Updated May 20 • 50 • 1
mradermacher/VeriReason-Qwen2.5-7b-SFT-Reasoning-GGUF Reinforcement Learning • 8B • Updated May 22 • 111 • 1
mradermacher/VeriReason-Qwen2.5-1.5B-grpo-small-GGUF Reinforcement Learning • 2B • Updated May 20 • 42 • 1
mradermacher/VeriReason-Qwen2.5-3B-Verilog-RTL-GRPO-reasoning-tb-GGUF Reinforcement Learning • 3B • Updated May 21 • 102
mradermacher/VeriReason-Qwen2.5-7b-SFT-Reasoning-i1-GGUF Reinforcement Learning • 8B • Updated May 22 • 208 • 1
mradermacher/VeriReason-Qwen2.5-1.5b-RTLCoder-Verilog-GRPO-reasoning-tb-GGUF Reinforcement Learning • 2B • Updated May 21 • 45
mradermacher/VeriReason-Qwen2.5-3b-RTLCoder-Verilog-GRPO-reasoning-tb-GGUF Reinforcement Learning • 3B • Updated May 21 • 58
mradermacher/VeriReason-Qwen2.5-7b-RTLCoder-Verilog-GRPO-reasoning-tb-GGUF Reinforcement Learning • 8B • Updated May 21 • 115 • 1
mradermacher/VeriReason-Qwen2.5-7b-RTLCoder-Verilog-GRPO-reasoning-tb-i1-GGUF Reinforcement Learning • 8B • Updated May 22 • 368 • 3