-
公开(公告)号:US11714681B2
公开(公告)日:2023-08-01
申请号:US16750100
申请日:2020-01-23
发明人: Hao Yang , Biswajit Das , Yu Gu , Peter Walker , Igor Karpenko , Robert Brian Christensen
CPC分类号: G06F9/5005 , G06F9/5027 , G06F9/5044 , G06F9/5055 , G06F9/5088 , G06F9/3836 , G06F9/3877 , G06N5/04 , G06T1/20
摘要: A method for dynamically assigning an inference request is disclosed. A method for dynamically assigning an inference request may include determining at least one model to process an inference request on a plurality of computing platforms, the plurality of computing platforms including at least one Central Processing Unit (CPU) and at least one Graphics Processing Unit (GPU), obtaining, with at least one processor, profile information of the at least one model, the profile information including measured characteristics of the at least one model, dynamically determining a selected computing platform from between the at least one CPU and the at least one GPU for responding to the inference request based on an optimized objective associated with a status of the computing platform and the profile information, and routing, with at least one processor, the inference request to the selected computing platform. A system and computer program product are also disclosed.
-
2.
公开(公告)号:US20230342203A1
公开(公告)日:2023-10-26
申请号:US18215921
申请日:2023-06-29
发明人: Hao Yang , Biswajit Das , Yu Gu , Peter Walker , Igor Karpenko , Robert Brian Christensen
IPC分类号: G06F9/50
CPC分类号: G06F9/5005 , G06F9/5027 , G06F9/5044 , G06F9/5088 , G06F9/5055 , G06F9/3836
摘要: A method for dynamically assigning an inference request is disclosed. A method for dynamically assigning an inference request may include determining at least one model to process an inference request on a plurality of computing platforms, the plurality of computing platforms including at least one Central Processing Unit (CPU) and at least one Graphics Processing Unit (GPU), obtaining, with at least one processor, profile information of the at least one model, the profile information including measured characteristics of the at least one model, dynamically determining a selected computing platform from between the at least one CPU and the at least one GPU for responding to the inference request based on an optimized objective associated with a status of the computing platform and the profile information, and routing, with at least one processor, the inference request to the selected computing platform. A system and computer program product are also disclosed.
-
公开(公告)号:US20210232399A1
公开(公告)日:2021-07-29
申请号:US16750100
申请日:2020-01-23
发明人: Hao Yang , Biswajit Das , Yu Gu , Peter Walker , Igor Karpenko , Robert Brian Christensen
摘要: A method for dynamically assigning an inference request is disclosed. A method for dynamically assigning an inference request may include determining at least one model to process an inference request on a plurality of computing platforms, the plurality of computing platforms including at least one Central Processing Unit (CPU) and at least one Graphics Processing Unit (GPU), obtaining, with at least one processor, profile information of the at least one model, the profile information including measured characteristics of the at least one model, dynamically determining a selected computing platform from between the at least one CPU and the at least one GPU for responding to the inference request based on an optimized objective associated with a status of the computing platform and the profile information, and routing, with at least one processor, the inference request to the selected computing platform. A system and computer program product are also disclosed.
-
公开(公告)号:US20210103927A1
公开(公告)日:2021-04-08
申请号:US16592374
申请日:2019-10-03
发明人: Navendu Misra , Biswajit Das , Durga Kala , Harshitha S V , Juharasha Shaik
IPC分类号: G06Q20/40
摘要: Methods, systems, and computer program products are provided for simulating fraud detection rules. The method includes: receiving user input including at least one proposed fraud detection rule in at least one input field of a user interface; simulating the at least one proposed fraud detection rule by: querying historical transaction data based on the at least one proposed fraud detection rule; and generating a simulation result for the at least one proposed fraud detection rule based on the queried historical transaction data; and displaying the simulation result on the user interface. The simulation result is displayed on the user interface asynchronously and in near real-time relative to receiving the user input.
-
-
-