> we ask them to stop testing advanced mathematical problems on proprietary models.
Maybe I'm alone on this, but for some reason these sorts of requests strike me as akin to gatekeeping how someone should breathe air. It's math... the numbers and symbols are just out there in the platonic realm available for anyone to do as they like with them. It's patently absurd to request other people to stop.
Ensuring credit where credit is due? That's fine. If your model incorporates the efforts of many others, then it's reasonable to request acknowledgement of everyone who contributed (even indirectly). But that's not what the request states — presumably their ask subsumes any advanced ML model, including those that weren't trained on a giant corpus of text.
I don't see why mathematicians should be protected from AI anymore than any other profession. It's either everybody or nobody, not fair on the face of it otherwise.
TFA isn't asking for mathematicians to be protected from AI. It's asking AI labs to hold themselves to the standards of the mathematical community:
- releasing papers using the normal process to allow peer review
- giving talks etc to disseminate knowledge so humans understand the result
- writing papers in a way (standard terminology etc) that allows mathematicians to digest the result (some AI math papers comprise a huge verbose load of non-standard terminology and waffle and then a massive lean proof. This is very hard for humans to actually understand, and means it's hard for others to take the work forward.)
- giving appropriate credit to results that are used to derive the work
It includes some specific recommendations for situations where the person prompting the model is not in a position to understand the output, and frankly these are really welcome given situations like the recent case at Anthropic where a non-mathematician at Anthropic prompted claude to make a significant improvement to the bounds of a problem related to the Riemann Zeta function[1] which led to widespread misreporting and claims (not by Anthropic themselves notably) that the Riemann hypothesis itself had been proved, which is emphatically not the case.
Research mathematics is fundamentally a collaborative activity and the way in which some of these results are released is done to maximise PR but means a ton of the mathematical value is left on the table.
[1] <a href="https://www.anthropic.com/research/riemann-zeta" rel="nofollow">https://www.anthropic.com/research/riemann-zeta. As I understand it, the Riemann Hypothesis says that all non-trivial zeroes of the zeta function lie on a line called the critical line. Two centuries of previous work had established that at least something like 40.9% of the zeroes lie on the line and noone has ever found a non-trivial zero that does not lie on that line. Claude (with prompting from a non-mathematician to "try harder" etc) improved this bound massively to 67%. Now a lot of people said things like "OK so all we've got to do is to improve that to 100% and we've proved the RH", which is definitely not true unfortunately, because you can say that in the limit the proportion of the zeroes on the line is 100% and still have infinitely many which are not.
Maintaining the existing standards is in fact a form of protection from AI disruption. They are asking the AI companies to follow their norms, instead of them having to conform to new norms created by AI. They don’t want to have to change the way they do things, understandably!, and are asking the companies to accommodate their way of life.
New norms are only worth adopting if they are clearly better, and that is far from obvious for whatever norms the proprietary AI companies are trying to push. Also their letter specifically targets proprietary AI companies, not computational tools in general which mathematicians do use when they advance mathematical understanding.
Better for who? If I don’t care about the welfare of mathematicians and just the advancement of mathematics, would AI doing the math not be better for me?
As well, if the open models were as good as the proprietary ones, do you think they wouldn’t still have complaints?
throwaway713 · · focus · HN ↗
Maybe I'm alone on this, but for some reason these sorts of requests strike me as akin to gatekeeping how someone should breathe air. It's math... the numbers and symbols are just out there in the platonic realm available for anyone to do as they like with them. It's patently absurd to request other people to stop.
Ensuring credit where credit is due? That's fine. If your model incorporates the efforts of many others, then it's reasonable to request acknowledgement of everyone who contributed (even indirectly). But that's not what the request states — presumably their ask subsumes any advanced ML model, including those that weren't trained on a giant corpus of text.
knuckleheads · · focus · HN ↗
seanhunter · · focus · HN ↗
- releasing papers using the normal process to allow peer review
- giving talks etc to disseminate knowledge so humans understand the result
- writing papers in a way (standard terminology etc) that allows mathematicians to digest the result (some AI math papers comprise a huge verbose load of non-standard terminology and waffle and then a massive lean proof. This is very hard for humans to actually understand, and means it's hard for others to take the work forward.)
- giving appropriate credit to results that are used to derive the work
It includes some specific recommendations for situations where the person prompting the model is not in a position to understand the output, and frankly these are really welcome given situations like the recent case at Anthropic where a non-mathematician at Anthropic prompted claude to make a significant improvement to the bounds of a problem related to the Riemann Zeta function[1] which led to widespread misreporting and claims (not by Anthropic themselves notably) that the Riemann hypothesis itself had been proved, which is emphatically not the case.
Research mathematics is fundamentally a collaborative activity and the way in which some of these results are released is done to maximise PR but means a ton of the mathematical value is left on the table.
[1] <a href="https://www.anthropic.com/research/riemann-zeta" rel="nofollow">https://www.anthropic.com/research/riemann-zeta. As I understand it, the Riemann Hypothesis says that all non-trivial zeroes of the zeta function lie on a line called the critical line. Two centuries of previous work had established that at least something like 40.9% of the zeroes lie on the line and noone has ever found a non-trivial zero that does not lie on that line. Claude (with prompting from a non-mathematician to "try harder" etc) improved this bound massively to 67%. Now a lot of people said things like "OK so all we've got to do is to improve that to 100% and we've proved the RH", which is definitely not true unfortunately, because you can say that in the limit the proportion of the zeroes on the line is 100% and still have infinitely many which are not.
knuckleheads · · focus · HN ↗
curt15 · · focus · HN ↗
knuckleheads · · focus · HN ↗
As well, if the open models were as good as the proprietary ones, do you think they wouldn’t still have complaints?