Needle: Finetune a 26M Tool-Calling Model Locally with Ollama @fahdmirza
Needle: Finetune a 26M Tool-Calling Model Locally with Ollama  @fahdmirza
Uploaded July 2026 | Updated September 2026, 2 weeks ago
This video installs and tests Needle which runs on Cactus at 6000 toks/sec prefill and 1200 decode speed.

🔥 Buy Me a Coffee to support the channel: ko-fi.com/fahdmirza

#cactusneedle #needleai

🔥 Get 50% Discount on any A6000 or A5000 GPU rental, use following link and coupon:

bit.ly/fahd-mirza
Coupon code: FahdMirza

🔥 Buy Me a Coffee to support the channel: ko-fi.com/fahdmirza

#needleai

PLEASE FOLLOW ME:
â–¶ LinkedIn: / fahdmirza
â–¶ YouTube: / @fahdmirza
â–¶ Blog: fahdmirza.com

RESOURCES:

â–¶ github.com/cactus-compute/needle

All rights reserved © Fahd Mirza
Needle: Finetune a 26M Tool-Calling Model Locally with OllamaIdeogram 4: Worlds Best Text-to-Image Model? Lets Test LocallyRun Alibaba OvisOCR2 Locally: First Model to Ever Beat the Pipeline-Based MethodsMaya1: Voice AI Model with Emotions! Run Locally for FreeBest Qwen3.6 Quant You Can Run Right Now LocallyI Built an Entire AI Media Empire With One Prompt (Skywork)MiniMax H3 Locally with ComfyUI and ClipProj on 1 GPUOpenDataLoader PDF: Open-Source PDF Parser for RAG Pipelines (Local, No GPU)SIE vs vLLM — Stop Using the Wrong Inference EngineGemini 3.6 Flash — Lost in the Australian OutbackARK-ASR-3B: Multilingual ASR Model Tested LocallyKimi K3 vs Fable 5 vs GLM 5.2 - An Unforgettable Showdown
Fahd Mirza |

Needle: Finetune a 26M Tool-Calling Model Locally with Ollama

SHARE TO X SHARE TO REDDIT SHARE TO FACEBOOK WALLPAPER