Running 26B and 35B LLMs at Full Speed on €990 of Used Hardware – No Cloud
Two large language models, of 26 billion and 35 billion parameters, can be run locally at full speed using only €990 worth of used hardware, avoiding the need for cloud services and demonstrating cost-effective access to powerful AI inference.