I’ve got access to an NVIDIA DGX Spark and want to see what open models can actually do on it beyond benchmark numbers.
Give me something worth testing: • a model you’re curious about • ridiculous context sizes • coding or agent workloads • vision/multimodal tasks • quantization comparisons • setups that seem like they shouldn’t work I’ll run some of the interesting ones and share what works, what struggles, and wha…
Read the full story at r/LocalLLM ↗