4 ms·
On many natural language tasks there can be significant overlap, making it difficult to judge performance. That's why I like more complex code generation tasks
by Inufu 4y ago
On many natural language tasks there can be significant overlap, making it difficult to judge performance.
That's why I like more complex code generation tasks such the dataset we used for AlphaCode.