ResearchDatasetsBIG-BenchFollowBenchmarksTask NameDataset NameSOTA MethodSortMost resultsRecently updatedMost papersApplyTask NameDataset NameSOTA ResultTrendResultsPerformance PredictionBIG-bench (train)2.487Boolean Expressions Score3General Language ModelingBIG-Bench (test)83.6Accuracy2Complex ReasoningBIG-bench Hard Orig QA—Original Metric Value0Page 3 of 3PreviousNext