Estimate BERT inference latency per token given hardware FLOPs, model layers, sequence length, and batch size.
Enter the values for the patient or scenario you are assessing. Estimate BERT inference latency per token given hardware FLOPs, model layers, sequence length, and batch size. Use the bert latency result to inform your clinical assessment.