Row 5206

Row ID: 5206 | Dataset Entry | Axioma AXP Content Repository

Content Data

This page contains data entry 5206 from the Axioma AXP content repository. The structured data below represents the complete record for this entry.

**Open Source Strikes Again**, We are thrilled to announce the release of OpenBioLLM-Llama3-70B & 8B. These models outperform industry giants like **Openai’s GPT-4, Google’s Gemini, Meditron-70B, Google’s Med-PaLM-1, and Med-PaLM-2** in the biomedical domain, setting a new state-of-the-art for models of their size. **The most capable openly available Medical-domain LLMs to date!** 🩺💊🧬

https://preview.redd.it/2h4ebhftf0xc1.png?width=2514&format=png&auto=webp&s=bbc3a583d45fb37b87a6fbbabe2d9e0f23c75d8b

🔥 OpenBioLLM-70B delivers SOTA performance, while the OpenBioLLM-8B model even surpasses GPT-3.5 and Meditron-70B!

The models underwent a rigorous two-phase fine-tuning process using the LLama-3 70B & 8B models as the base and leveraging Direct Preference Optimization (DPO) for optimal performance. 🧠

https://preview.redd.it/w41pv7mwf0xc1.png?width=5760&format=png&auto=webp&s=f3143919ef8472961f329bb8eb98937d8f8e41e0

**Results are available at Open Medical-L**LM Leaderboard: [https://huggingface.co/spaces/openlifescienceai/open\_medical\_llm\_leaderboard](https://huggingface.co/spaces/openlifescienceai/open_medical_llm_leaderboard)

Over \~4 months, we meticulously curated a diverse custom dataset, collaborating with medical experts to ensure the highest quality. The dataset spans 3k healthcare topics and 10+ medical subjects. 📚 OpenBioLLM-70B's remarkable performance is evident across 9 diverse biomedical datasets, achieving an impressive average score of 86.06% despite its smaller parameter count compared to GPT-4 & Med-PaLM. 📈

https://preview.redd.it/5ff2k9szf0xc1.png?width=5040&format=png&auto=webp&s=15dc4aa948f2608717f68ddf2cb27a6a2de03496

You can download the models directly from Huggingface today.

- 70B : [https://huggingface.co/aaditya/OpenBioLLM-Llama3-70B](https://huggingface.co/aaditya/OpenBioLLM-Llama3-70B) - 8B : [https://huggingface.co/aaditya/OpenBioLLM-Llama3-8B](https://huggingface.co/aaditya/OpenBioLLM-Llama3-8B)

This release is just the beginning! In the coming months, we'll introduce

* Expanded medical domain coverage, * Longer context windows, * Better benchmarks, and * Multimodal capabilities.

More details can be found here: [https://twitter.com/aadityaura/status/1783662626901528803](https://twitter.com/aadityaura/status/1783662626901528803) Over the next few months, Multimodal will be made available for various medical and legal benchmarks.

I hope it's useful in your research 🔬 Have a wonderful weekend, everyone! 😊

FieldValue
text **Open Source Strikes Again**, We are thrilled to announce the release of OpenBioLLM-Llama3-70B & 8B. These models outperform industry giants like **Openai’s GPT-4, Google’s Gemini, Meditron-70B, Google’s Med-PaLM-1, and Med-PaLM-2** in the biomedical domain, setting a new state-of-the-art for models of their size. **The most capable openly available Medical-domain LLMs to date!** 🩺💊🧬 https://preview.redd.it/2h4ebhftf0xc1.png?width=2514&format=png&auto=webp&s=bbc3a583d45fb37b87a6fbbabe2d9e…
label r/machinelearning
dataType post
communityName r/MachineLearning
datetime 2024-04-27
username_encoded Z0FBQUFBQm5LakwyRkRHclA1U3RKN2Noem1MWXRhY21xeG56cnM2b3hwdEdCTmVuT0JpWXcxcjY2dnhNV096T2VoT0lLUHV5QnIzZWl2RTRxOXotR3VBc25UZTA0REhzSmc9PQ==
url_encoded Z0FBQUFBQm5Lak9GTGNvazNzYl92MmNVREJhV3JFZGtMM1VyTW9zMG1Ma1dvTkNtRDVfaUhrNjJOLU92czlGQkd5Ry1yRVdBNnR2TmdwWjRpR2pKUXdpZDBEU3FCMWJFRzFQMTdsOXU0bklnSFloSUx4YWJsUjBRd2VFdTNRTXZwVHM0YVVac0oydTljSi02QkFzQy1zN0RmUEQ1d3ZEeDFhcFNBVmEwOUFnaDMxWTZuZkQxSFhaS2dRRC1jTEptejNHa1RaNXA2U2Nva1hWcHo4RlQzRlNIU0RCRGptS3JzUT09

Raw Record

{
  "text": "**Open Source Strikes Again**, We are thrilled to announce the release of OpenBioLLM-Llama3-70B & 8B. These models outperform industry giants like **Openai’s GPT-4, Google’s Gemini, Meditron-70B, Google’s Med-PaLM-1, and Med-PaLM-2** in the biomedical domain, setting a new state-of-the-art for models of their size. **The most capable openly available Medical-domain LLMs to date!** 🩺💊🧬\n\n\n\nhttps://preview.redd.it/2h4ebhftf0xc1.png?width=2514&format=png&auto=webp&s=bbc3a583d45fb37b87a6fbbabe2d9e0f23c75d8b\n\n🔥 OpenBioLLM-70B delivers SOTA performance, while the OpenBioLLM-8B model even surpasses GPT-3.5 and Meditron-70B!\n\nThe models underwent a rigorous two-phase fine-tuning process using the LLama-3 70B & 8B models as the base and leveraging Direct Preference Optimization (DPO) for optimal performance. 🧠\n\n\n\nhttps://preview.redd.it/w41pv7mwf0xc1.png?width=5760&format=png&auto=webp&s=f3143919ef8472961f329bb8eb98937d8f8e41e0\n\n**Results are available at Open Medical-L**LM Leaderboard: [https://huggingface.co/spaces/openlifescienceai/open\\_medical\\_llm\\_leaderboard](https://huggingface.co/spaces/openlifescienceai/open_medical_llm_leaderboard)\n\nOver \\~4 months, we meticulously curated a diverse custom dataset, collaborating with medical experts to ensure the highest quality. The dataset spans 3k healthcare topics and 10+ medical subjects. 📚 OpenBioLLM-70B's remarkable performance is evident across 9 diverse biomedical datasets, achieving an impressive average score of 86.06% despite its smaller parameter count compared to GPT-4 & Med-PaLM. 📈\n\n\n\nhttps://preview.redd.it/5ff2k9szf0xc1.png?width=5040&format=png&auto=webp&s=15dc4aa948f2608717f68ddf2cb27a6a2de03496\n\nYou can download the models directly from Huggingface today.\n\n- 70B : [https://huggingface.co/aaditya/OpenBioLLM-Llama3-70B](https://huggingface.co/aaditya/OpenBioLLM-Llama3-70B)  \n- 8B : [https://huggingface.co/aaditya/OpenBioLLM-Llama3-8B](https://huggingface.co/aaditya/OpenBioLLM-Llama3-8B)\n\nThis release is just the beginning! In the coming months, we'll introduce\n\n* Expanded medical domain coverage,\n* Longer context windows,\n* Better benchmarks, and\n* Multimodal capabilities.\n\nMore details can be found here: [https://twitter.com/aadityaura/status/1783662626901528803](https://twitter.com/aadityaura/status/1783662626901528803) Over the next few months, Multimodal will be made available for various medical and legal benchmarks.\n\nI hope it's useful in your research 🔬 Have a wonderful weekend, everyone! 😊",
  "label": "r/machinelearning",
  "dataType": "post",
  "communityName": "r/MachineLearning",
  "datetime": "2024-04-27",
  "username_encoded": "Z0FBQUFBQm5LakwyRkRHclA1U3RKN2Noem1MWXRhY21xeG56cnM2b3hwdEdCTmVuT0JpWXcxcjY2dnhNV096T2VoT0lLUHV5QnIzZWl2RTRxOXotR3VBc25UZTA0REhzSmc9PQ==",
  "url_encoded": "Z0FBQUFBQm5Lak9GTGNvazNzYl92MmNVREJhV3JFZGtMM1VyTW9zMG1Ma1dvTkNtRDVfaUhrNjJOLU92czlGQkd5Ry1yRVdBNnR2TmdwWjRpR2pKUXdpZDBEU3FCMWJFRzFQMTdsOXU0bklnSFloSUx4YWJsUjBRd2VFdTNRTXZwVHM0YVVac0oydTljSi02QkFzQy1zN0RmUEQ1d3ZEeDFhcFNBVmEwOUFnaDMxWTZuZkQxSFhaS2dRRC1jTEptejNHa1RaNXA2U2Nva1hWcHo4RlQzRlNIU0RCRGptS3JzUT09"
}

Entry Information