local AI inference
5 stories tagged local AI inference.
Intel Arc Pro B70: 32GB of VRAM at a Real Price
Intel's Arc Pro B70 offers 32GB of VRAM for $1,000 and free SR-IOV support. For homelabbers and AI tinkerers, it's a serious buy. Gamers should wait.
Diffusion Gemma Runs Locally—and That Changes Privacy
Diffusion Gemma Runs Locally—and That Changes Privacy
Google's Diffusion Gemma runs on consumer GPUs at 700+ tokens/sec. For privacy, the real story isn't speed—it's that your prompts never leave your machine.
Llama.cpp Gets MTP: Local AI Just Got Faster
Llama.cpp Gets MTP: Local AI Just Got Faster
Llama.cpp just merged Multi-Token Prediction, giving local AI a ~25% speed boost. Here's why that matters for your privacy—and how to use it.
Desktop AI Supercomputers: What Dell's GB10 Says About Tech
Desktop AI Supercomputers: What Dell's GB10 Says About Tech
Dell's Pro Max with GB10 brings Nvidia's Blackwell chips to your desk. But who needs a 1 petaflop AI workstation at home, and what does it signal about computing's future?
Intel Arc Pro B60: Testing 96GB of AI VRAM for $5K
Intel Arc Pro B60: Testing 96GB of AI VRAM for $5K
Level1Techs tests Intel's Battle Matrix with four Arc Pro B60 GPUs—96GB VRAM for the price of an RTX 5090. Real-world AI performance examined.