텐서플로 서빙에 대해 배치 예측을 활성화한 다음 사용 사례에 맞게 구성해야 합니다. 다음과
같은 다섯 가지 구성 옵션이 있습니다.
●
max_batch_size
배치 크기를 제어하는 매개변수입니다. 배치 크기가 크면 요청 지연 시간이 증가하고
GPU
메모리가 소
진될 수 있습니다. 배치 크기가 작으면 최적의 계산 리소스를 사용할 수 있는 이점이 없어집니다.
●
batch_timeout_micros
배치를 채우기 위한 최대 대기 시간을 설정하는 매개변수입니다. 추론 요청에 대한 지연 시간을 제한하
는 데 유용합니다.
●
num_batch_threads
스레드 수는 병렬로 사용할 수 있는
CPU
나
GPU
코어 수를 구성합니다.
226
살아 움직이는 머신러닝 파이프라인 설계
●
max_enqueued_batches ...
Become an O’Reilly member and get unlimited access to this title plus top books and audiobooks from O’Reilly and nearly 200 top publishers, thousands of courses curated by job role, 150+ live events each month, and much more.
O’Reilly covers everything we've got, with content to help us build a world-class technology community, upgrade the capabilities and competencies of our teams, and improve overall team performance as well as their engagement.
Julian F.
Head of Cybersecurity
I wanted to learn C and C++, but it didn't click for me until I picked up an O'Reilly book. When I went on the O’Reilly platform, I was astonished to find all the books there, plus live events and sandboxes so you could play around with the technology.
Addison B.
Field Engineer
I’ve been on the O’Reilly platform for more than eight years. I use a couple of learning platforms, but I'm on O'Reilly more than anybody else. When you're there, you start learning. I'm never disappointed.
Amir M.
Data Platform Tech Lead
I'm always learning. So when I got on to O'Reilly, I was like a kid in a candy store. There are playlists. There are answers. There's on-demand training. It's worth its weight in gold, in terms of what it allows me to do.