Gcore Unveils Inference at the Edge – Bringing AI Applications Closer to End Users for Seamless Real-Time Performance
6.6.2024 11:30:00 EEST | Business Wire | Press release
Gcore, the global edge AI, cloud, network, and security solutions provider, today announced the launch of Gcore Inference at the Edge, a breakthrough solution that provides ultra-low latency experiences for AI applications. This innovative solution enables the distributed deployment of pre-trained machine learning (ML) models to edge inference nodes, ensuring seamless, real-time inference.
This press release features multimedia. View the full release here: https://www.businesswire.com/news/home/20240606719181/en/
Gcore Inference at the Edge empowers businesses across diverse industries with cost-effective, scalable, and secure AI model deployment (Graphic: Gcore)
Gcore Inference at the Edge empowers businesses across diverse industries—including automotive, manufacturing, retail, and technology—with cost-effective, scalable, and secure AI model deployment. Use cases such as generative AI, object recognition, real-time behavioural analysis, virtual assistants, and production monitoring can now be rapidly realised on a global scale.
Gcore Inference at the Edge runs on Gcore's extensive global network of 180+ edge nodes, all interconnected by Gcore's sophisticated low-latency smart routing technology. Each high-performance node sits at the edge of the Gcore network, strategically placing servers close to end users. Inference at the Edge runs on NVIDIA L40S GPUs, the market-leading chip designed specifically for AI inference. When a user sends a request, an edge node determines the route to the nearest available inference region with the lowest latency, achieving a typical response time of under 30 ms.
The new solution supports a wide range of fundamental ML and custom models. Available open-source foundation models in the Gcore ML Model Hub include LLaMA Pro 8B, Mistral 7B, and Stable-Diffusion XL. Models can be selected and trained agnostically to suit any use case, before distributing them globally to Gcore Inference at the Edge nodes. This addresses a significant challenge faced by development teams where AI models are typically run on the same servers they were trained on, resulting in poor performance.
Benefits of Gcore Inference at the Edge include:
- Cost-effective deployment: A flexible pricing structure ensures customers only pay for the resources they use.
- Inbuilt DDoS protection: ML endpoints are automatically protected from DDoS attacks through Gcore’s infrastructure.
- Outstanding data privacy and security: The solution features built-in compliance with GDPR, PCI DSS, and ISO/IEC 27001 standards.
- Model autoscaling: Autoscaling is available to handle load spikes, so a model is always ready to support peak demand and unexpected surges.
- Unlimited object storage: Scalable S3-compatible cloud storage that grows with evolving model needs.
Andre Reitenbach, CEO at Gcore comments: “Gcore Inference at the Edge empowers customers to focus on getting their machine learning models trained, rather than worrying about the costs, skills, and infrastructure required to deploy AI applications globally. At Gcore, we believe the edge is where the best performance and end-user experiences are achieved, and that is why we are continuously innovating to ensure every customer receives unparalleled scale and performance. Gcore Inference at the Edge delivers all the power with none of the headache, providing a modern, effective, and efficient AI inference experience.”
Learn more at https://gcore.com/inference-at-the-edge
About Gcore
Gcore is the global edge AI, cloud, network, and security solutions provider. Gcore provides its solutions to global leaders in numerous industries. The company manages its own global IT infrastructure across six continents, with one of the best network performances in Europe, Africa, and LATAM, due to the average response time of 30 ms worldwide. Gcore’s network consists of 180+ points of presence around the world in reliable Tier IV and Tier III data centres, with a total capacity exceeding 200 Tbps.
To view this piece of content from cts.businesswire.com, please give your consent at the top of this page.
View source version on businesswire.com: https://www.businesswire.com/news/home/20240606719181/en/
Contact information
Gcore press contact
pr@gcore.com
About Business Wire
For more than 50 years, Business Wire has been the global leader in press release distribution and regulatory disclosure.
Subscribe to releases from Business Wire
Subscribe to all the latest releases from Business Wire by registering your e-mail address below. You can unsubscribe at any time.
Latest releases from Business Wire
NetApp and Oracle to Launch Fully Managed Cloud Storage Service for Demanding AI and Enterprise Workloads29.9.2026 20:30:00 EEST | Press release
NetApp® (NASDAQ: NTAP), the Intelligent Data Infrastructure company, and Oracle today announced a new fully managed storage service that brings enterprise-grade NetApp storage natively to Oracle Cloud Infrastructure (OCI). The service will bring NetApp ONTAP® data management capabilities natively to OCI to help organizations simplify the migration and management of critical AI and enterprise workloads in the cloud. Oracle Cloud Infrastructure NetApp Storage Service is designed to be a fully managed OCI-native storage service that brings NetApp ONTAP data management capabilities to OCI. It helps enterprises migrate, run, protect, and modernize databases, enterprise applications, virtualized environments, EDA/HPC, regulated applications, and AI data pipelines while preserving familiar storage operations. “Customers want the flexibility to move AI and enterprise workloads to the cloud without giving up operational simplicity, data management, and reliability,” said Pravjit Tiwana, Senior
NetApp to Expand Partnership with SAP to Support SAP Cloud Infrastructure29.9.2026 20:30:00 EEST | Press release
NetApp, the intelligent Data Infrastructure company, today announced its intent to expand its long-standing partnership with SAP, with plans to explore deeper cloud-native integration of NetApp capabilities into SAP Cloud Infrastructure, SAP’s Infrastructure-as-a-Service platform. The NetApp Platform supports key elements of the SAP Cloud Infrastructure storage architecture, delivering reliable, highly available and high-performance file and block storage services. The NetApp Platform also provides built-in proactive data protection capabilities and data replication to support SAP’s global availability model. “SAP and NetApp share a long-standing commitment to empower the world’s most mission-critical enterprise environments,” said Álvaro Celis, Senior Vice President, Chief Commercial, Partner and Ecosystem Officer at NetApp. “By bringing together SAP Cloud Infrastructure with the NetApp Platform, we can empower customers to build an intelligent data infrastructure that delivers the ag
NetApp Honors the International Union for Conservation of Nature with the Social Impact Award for Intelligent Data for Good29.9.2026 20:30:00 EEST | Press release
NetApp® (NASDAQ: NTAP), the Intelligent Data Infrastructure company, today announced that it awarded the International Union for Conservation of Nature (IUCN) the 2026 Social Impact Award for Intelligent Data for Good at NetApp INSIGHT 2026, NetApp's flagship global event dedicated to technological innovation, taking place in Las Vegas from September 29 to October 1. Part of NetApp’s Social Impact Awards program, this recognition honors customers and partners using NetApp technology in support of social or environmental impact initiatives. The Intelligent Data for Good category reflects NetApp’s commitment to recognizing and supporting organizations using intelligent data to improve lives and strengthen communities. IUCN is the world's largest and most diverse environmental network, bringing together governments, civil society organizations, and experts to advance conservation action worldwide. Uniquely composed of more than 1,600 government and civil society member organizations and 1
NetApp and Supermicro Collaborate to Power AI at Any Scale29.9.2026 20:30:00 EEST | Press release
NetApp® (NASDAQ: NTAP), the Intelligent Data Infrastructure company, today announced a strategic collaboration with Supermicro to deliver jointly validated AI infrastructure solutions spanning large-scale AI factories, enterprise deployments, neoclouds, and sovereign AI initiatives. By combining Supermicro's AI-optimized compute and rack-scale systems with the NetApp Platform, the companies will help customers accelerate deployment, improve GPU utilization, strengthen data governance, and reduce operational complexity. NetApp and Supermicro are creating a full-stack foundation for AI that extends beyond high-performance storage. The collaboration brings together Supermicro’s rack-scale compute, networking, power, and liquid-cooling expertise with the NetApp Platform’s unified data management, cyber resilience, hybrid-cloud leadership, and disaggregated scale. Customers gain an architecture that can be deployed quickly, evolve across generations of compute, and deliver a consistent data
NetApp Removes Storage Bottleneck for AI Factories29.9.2026 20:30:00 EEST | Press release
NetApp® (NASDAQ: NTAP), the Intelligent Data Infrastructure company, today announced NetApp Novus™, the fastest storage on the planet designed for AI infrastructure providers operating large GPU environments. Built to support a zettabyte-scale file system, this next-generation architecture will allow AI factories to keep millions of GPUs continuously fed with data, maximizing the efficiency, utilization, and profitability of their investments. Neoclouds and GPU-as-a-Service providers are building AI factories at a scale that traditional storage architectures were never designed to support, creating bottlenecks that reduce GPU utilization and increase infrastructure costs. Under traditional architectures, GPU utilization can drop below 30 percent when AI factories cannot feed enough data to their GPUs to generate the results they need and make a return on their investment. “AI factories struggle and GPU economics collapse when data can't keep up," said Syam Nair, Chief Product Officer a
In our pressroom you can read all our latest releases, find our press contacts, images, documents and other relevant information about us.
Visit our pressroom
