Hook

Their other posts in the index, biggest breakout first.
Scientists tested a 3.5 billion dollar medical AI against regular chat GPT and it lost on every single question. There's this AI tool called Open Evidence, built specifically for doctors. It just raised $210 billion and hospitals are buying and using it. Another one is up to date. Doctors have used it for years. These tools assume specialized medical AI is better than regular chat GPT. But nobody had tested this until NYU. They ran studies and the last one matters most. They took 100 questions actual doctors typed into AI while treating patients. NYU ran them through medical tools like open evidence and up to date and also through GPT, Cloud and Gemini. Then 12 doctors graded every answer without knowing which AI wrote it. Those expensive medical tools they lost. Every doctor ranked GPT, Cloud and Gemini above the medical specific ones. These medical tools pull from medical databases using RA. But research shows RA can make answers worse if wrong information is pulled versus Frontier AI which has answers baked into its This matters because these tools are already in hospitals. Doctors are making real decisions with them. The paper highlights. Nobody independently tested any of this before. No one asked if a $700 medical AI is better than a free one before using it on patients. Stop assuming specialized means better. These A's have great marketing until tested.