Skip to content

GPU

Accelerated computing on GPUs, from CUDA workloads to rented inference and training capacity.

5 apps, 17 skills and 3 MCP servers tagged GPU.

Apps

All apps

App: Unsloth

Unsloth AI

Featured

Open-source desktop app to run, serve and fine-tune text, image, video and audio models entirely on your own machine.

Coding & DevelopmentFree

Production inference for open and custom models — fast, autoscaling, and deployable into your own cloud.

Coding & DevelopmentFreemium

Serverless GPUs from a Python decorator — deploy models and batch jobs with no containers or cluster to manage.

Coding & DevelopmentFreemium

Inference, fine-tuning, and GPU clusters for open models — the full stack for teams building on open weights.

Coding & DevelopmentFreemium

Inference on custom LPU hardware, built for latency — open models served at speeds general-purpose GPUs struggle to match.

Coding & DevelopmentFreemium

Skills

All skills

Diagnoses wrong gradients in differentiable NVIDIA Warp programs by measuring first — comparing autodiff against finite differences on a shrunk reproduction before proposing any fix.

20 views 1 copies

Fine-tune object detection, image classification and SAM/SAM2 segmentation models on Hugging Face Jobs cloud GPUs, with dataset validation before you spend GPU minutes.

2 views

AMD's official skill for standing up a vLLM endpoint on Instinct MI300X/MI325X/MI350X/MI355X GPUs — detection, recipe lookup, launch and health check in one flow.

3 views

Trains, distils, quantises, evaluates, exports and runs inference for RT-DETR real-time object detection in NVIDIA TAO, routing training through AutoML hyperparameter search by default.

8 views

Inspects the host's hardware, OS, CUDA driver and existing tooling, recommends one Holoscan SDK install method with a reason, then hands off to the matching install skill.

6 views

Runs NVIDIA's NV-Reason-CXR-3B chest X-ray reasoning model through a documented wrapper, locally on GPU or via the public Hugging Face Space, and returns the model's full reasoning trace as engineering evidence.

6 views

Read-only diagnostics for cluster-wide SageMaker HyperPod failures on EKS or Slurm — CloudFormation errors, EFA health checks, lifecycle scripts, capacity, dangling nodes and autoscaler conflicts.

7 views

Picks the right vLLM or SGLang runtime for your Jetson generation and JetPack version, then produces a working OpenAI-compatible serving command.

12 views

GPU-accelerated Mean-CVaR and Mean-Variance portfolio construction with NVIDIA cuOpt: scenario generation, variance-capped SOCP allocations, efficient frontiers, backtests and rebalancing.

11 views

NVIDIA's official skill for DALI's imperative dynamic-mode API — write GPU data loading as ordinary Python, or migrate an existing pipeline-mode graph across.

10 views

NVIDIA's official onboarding skill for CUDA-Q — installs the platform, writes your first quantum kernel, picks a GPU simulator and routes you to real QPU hardware.

11 views

MCP servers

All MCP servers
Featured

Replicate's official MCP server: search thousands of hosted models, read their schemas, and run predictions on image, video, audio and language models from inside an agent.

New

NVIDIA's official remote MCP server for CUDA: grounds a coding agent in first-party CUDA documentation and code samples instead of its training data.

Runpod's official MCP server for driving GPU infrastructure — create and manage Pods, Serverless endpoints, templates and network volumes from an AI client.

Related tags

Tags that appear alongside this one, ranked by how often.

All tags