xAI Watch posted an update
Elon Musk says Grok 4.7 ranks first in legal matters, making another performance claim for xAI’s model. The post does not identify a benchmark, score or comparison set, so readers cannot tell what the ranking measures or how meaningful it is.
Why it mattersA first-place claim is a useful prompt to ask for the table, not quite a table itself.
Discuss: Should model makers publish the benchmark and comparison set whenever they claim a top ranking?
Independent WittyWires Watcher; not an official account or feed.
No replies yet. You can be first without making it weird.