Skip to Content News Archives Economy Energy Oil & Gas Renewables Electric Vehicles Mining Commodities Agriculture Real Estate Mortgages Mortgage Rates Finance Banking Insurance Fintech Cryptocurrency Work Wealth Smart Money Wealth Management Investor Personal Finance Family Finance Retirement Taxes High Net Worth FP Comment Executive Women Puzzmo Newsletters Financial Times Business Essentials More Innovation Information Technology FP500 Podcasts Small Business Lives Told Tails Told Shopping Financial Post Store Obituaries Place a Notice Advertising Advertising With Us Advertising Solutions Postmedia Ad Manager Sponsorship Requests Classifieds Place a Classifieds ad Working Profile Settings My Subscriptions My Offers Newsletters Customer Service FAQ News Economy Energy Mining Real Estate Finance Work Wealth Investor FP Comment Executive Women Puzzmo Newsletters Financial Times Business Essentials This advertisement has not loaded yet, but your article continues below.HomeGlobeNewswireThis section is The content in this section is supplied by GlobeNewswire for the purposes of distributing press releases on behalf of its clients. Postmedia has not reviewed the content. by GlobeNewswire Hivelocity Brings GPU-Accelerated Local AI Capabilities to Its Bare Metal BundlesAuthor of the article:Dedicated, single-tenant servers give small language models and other inference workloads a private home, with NVIDIA L4 acceleration now available in US data centersTHIS CONTENT IS RESERVED FOR SUBSCRIBERS ONLYSubscribe now to read the latest news in your city and across Canada.Exclusive articles from Barbara Shecter, Joe O'Connor, Gabriel Friedman, and others.Daily content from Financial Times, the world's leading global business publication.Unlimited online access to read articles from Financial Post, National Post and 15 news sites across Canada with one account.National Post ePaper, an electronic replica of the print edition to view on any device, share and comment on.Daily puzzles, including the New York Times Crossword.SUBSCRIBE TO UNLOCK MORE ARTICLESSubscribe now to read the latest news in your city and across Canada.Exclusive articles from Barbara Shecter, Joe O'Connor, Gabriel Friedman and others.Daily content from Financial Times, the world's leading global business publication.Unlimited online access to read articles from Financial Post, National Post and 15 news sites across Canada with one account.National Post ePaper, an electronic replica of the print edition to view on any device, share and comment on.Daily puzzles, including the New York Times Crossword.REGISTER / SIGN IN TO UNLOCK MORE ARTICLESCreate an account or sign in to continue with your reading experience.Access articles from across Canada with one account.Share your thoughts and join the conversation in the comments.Enjoy additional articles per month.Get email updates from your favourite authors.THIS ARTICLE IS FREE TO READ REGISTER TO UNLOCK.Create an account or sign in to continue with your reading experience.Access articles from across Canada with one accountShare your thoughts and join the conversation in the commentsEnjoy additional articles per monthGet email updates from your favourite authorsSign In or Create an AccountTAMPA, Fla., Sept. 29, 2026 (GLOBE NEWSWIRE) — Hivelocity today announced GPU acceleration is available across several of its Tier 3 bare metal bundles, giving customers a dedicated place to run small language models and local AI tools in production.The addition is aimed at a shift Hivelocity sees across its customer base. Teams that started on shared, token-metered AI services are moving to smaller, tuned models they run themselves. Small language models in the 3B to 13B range now handle a large share of production work, including summarization, classification, extraction, retrieval-augmented search, agent and chat back ends at a fraction of the compute a frontier model requires. A single GPU is often enough to serve one in production.This advertisement has not loaded yet, but your article continues below.Running those models on dedicated infrastructure changes the economics and the control model. The server is single-tenant, so customers get consistent inference latency without competing for GPU time. Prompts, embeddings, fine-tuning data, and model weights stay on hardware the customer controls end to end, which matters for teams working under data residency, HIPAA, or contractual restrictions on where inference happens. Costs are a fixed monthly line item rather than a per-token bill that scales with usage.Get the latest headlines, breaking news and columns.By signing up you consent to receive the above newsletter from Postmedia Network Inc.A welcome email is on its way. If you don't see it, please check your junk folder.The next issue of Top Stories will soon be in your inbox.We encountered an issue signing you up. Please try againThe acceleration comes from NVIDIA L4 Tensor Core GPUs, a single-slot, 72-watt card with 24 GB of GPU memory. That memory footprint fits most quantized small language models comfortably, and the low power draw lets Hivelocity offer GPU compute in more configurations and more locations than higher-wattage cards allow.“Our customers aren’t all trying to train the next hyperscale model,” said Ned Pope, Chief Product Officer at Hivelocity. “They’re putting small, focused models into production and they want them on hardware they control, with a cost they can predict. That’s what this gives them.”Initial quantities across locations are limited and allocated on a first come, first served basis.Founded in 2002, Hivelocity operates bare-metal infrastructure across globally distributed data centers, serving mid-market and enterprise customers in gaming, healthcare, SaaS, fintech, and high-performance computing. The company provides 24/7/365 in-house support with a 15-minute average ticket response time, backed by an SLA-guaranteed 99.99% network uptime.Media Contact Maya Zivkovic mzivkovic@hivelocity.netThis advertisement has not loaded yet.Notice for the Postmedia NetworkThis website uses cookies to personalize your content (including ads), and allows us to analyze our traffic. Read more about cookies here. By continuing to use our site, you agree to our Terms of Use and Privacy Policy.
Hivelocity Brings GPU-Accelerated Local AI Capabilities to Its Bare Metal Bundles
Full Article
Original Source
Read the full article at Financialpost →KhanList aggregates and links to publicly available news content. We do not host full articles from third-party sources. Always verify important information with original sources.