AI & ML interests

AI cybersecurity serves as the guardian of our digital world, protecting against malicious threats and hackers, including those who use artificial intelligence.

legolasyiuย 
posted an update about 2 months ago
view post
Post
204
Introducing Reasoning-Medical-27B is designed for advanced medical reasoning in professional medicine, medical genetics, college biology/medicine, and clinical knowledge. The model was fine-tuned on a large-scale dataset of 370,000 high-quality question-and-answer examples, incorporating Chain-of-Thought reasoning to improve step-by-step problem solving. Training was performed using the GRPO trainer with the Unsloth optimization method for efficient fine-tuning.
MedQA: 93% vs MedGemma 85.3%

Model: EpistemeAI/Reasoning-Medical-27B



# Benchmark

| Task              | Version | Filter         | n-shot | Metric      | Direction | Reasoning Medical 27B | Qwen 3.6 27B | MedGemma 1 27B |
|-------------------|--------:|----------------|-------:|-------------|:---------:|----------------------:|-------------:|----------------:|
| MMLU-Pro Biology  | 3.1     | custom-extract | 2      | exact_match | โ†‘         | 0.85                  | โ€”            | โ€”               |
| MMLU-ProX Biology | 0       | custom-extract | 2      | exact_match | โ†‘         | 0.80                  | โ€”            | โ€”               |
| MedQA             | YAML    | none           | 2      | acc         | โ†‘         | 0.93                  | 0.844        | 0.853           |

  • 4 replies
ยท
legolasyiuย 
posted an update 8 months ago
view post
Post
375
We release open-weight early experimental Codeforce metatune-gpt20b, fine tuned version of OpenAI's gpt-oss-20b model, this is one of the first public release recursive self improving AI.

EpistemeAI/Codeforce-metatune-gpt20b
  • 9 replies
ยท
legolasyiuย 
updated a Space over 1 year ago
legolasyiuย 
published a Space over 1 year ago