Inference is the execution of a trained artificial intelligence or machine learning model on new data to produce outputs such as predictions or classifications, used in enterprises to support production applications, automated decisions, and analytics within governed, monitored environments.
Redis announced Redis Feature Form, a managed enterprise feature store platform for defining, orchestrating, versioning, and serving ML features across training and inference. The release adds multi-tenant workspaces, fine-grained job control, atomic DAG updates, enhanced RBAC/security, simplified deployment, and a redesigned dashboard.
SK hynix began mass production of 192GB SOCAMM2, an LPDDR5X low-power DRAM server memory module built on its 1cnm process. The company claims over double bandwidth and over 75% improved power efficiency versus RDIMM, and positions the module for NVIDIA Vera Rubin–based next-generation AI servers.
NVIDIA released Dynamo 1.0, an open source software designed for AI inference at scale, integrating with frameworks and supported by major cloud providers and enterprises.