You have recently deployed a model to a Vertex AI endpoint and configured online serving in Vertex AI Feature Store. You have set up a daily batch-ingestion job to update your featurestore. During the batch-ingestion jobs, you find that CPU utilization is high on your featurestore’s online-serving nodes and that feature-retrieval latency is high. You need to improve online-serving performance during the daily batch ingestion. What should you do?
Community Discussion
No comments yet. Be the first to start the discussion!
Community Discussion