Add link to paper

#10
by nielsr HF Staff - opened
Files changed (1) hide show
  1. README.md +6 -4
README.md CHANGED
@@ -1,12 +1,11 @@
1
  ---
2
- model_name: PhAI-IDE-72B
3
  base_model: Qwen/Qwen2.5-72B-Instruct
4
- base_model_relation: finetune
5
  library_name: transformers
6
- pipeline_tag: text-generation
7
  license: other
8
  license_name: qwen
9
  license_link: https://huggingface.co/Qwen/Qwen2.5-72B-Instruct/blob/495f39366efef23836d0cfae4fbe635880d2be31/LICENSE
 
 
10
  tags:
11
  - phai-ide
12
  - science
@@ -15,10 +14,13 @@ tags:
15
  - sft
16
  - lora
17
  - safetensors
 
18
  ---
19
 
20
  # PhAI-IDE
21
 
 
 
22
  **PhAI-IDE** is a family of models for scientific coding and interaction with tools, available in **4B, 9B, and 72B** sizes. Each model is supervised fine-tuned with [ms-swift](https://github.com/modelscope/ms-swift) and released as full BF16 weights with the final LoRA adapter merged, together with its configuration and tokenizer.
23
 
24
  The **training dataset is Codex trajectories**, sourced from [ScienceIDE](https://github.com/aitofound/ScienceIDE).
@@ -109,4 +111,4 @@ The release was validated with the following environment.
109
  | PEFT | 0.20.0 |
110
  | Datasets | 4.8.4 |
111
  | Tokenizers | 0.23.2 |
112
- | Accelerate | 1.14.0 |
 
1
  ---
 
2
  base_model: Qwen/Qwen2.5-72B-Instruct
 
3
  library_name: transformers
 
4
  license: other
5
  license_name: qwen
6
  license_link: https://huggingface.co/Qwen/Qwen2.5-72B-Instruct/blob/495f39366efef23836d0cfae4fbe635880d2be31/LICENSE
7
+ model_name: PhAI-IDE-72B
8
+ pipeline_tag: text-generation
9
  tags:
10
  - phai-ide
11
  - science
 
14
  - sft
15
  - lora
16
  - safetensors
17
+ base_model_relation: finetune
18
  ---
19
 
20
  # PhAI-IDE
21
 
22
+ This repository contains the **PhAI-IDE-72B** model presented in [ScienceIDE: Turning World's Scientific Codebase into Agent Learnable Environments](https://huggingface.co/papers/2609.19134).
23
+
24
  **PhAI-IDE** is a family of models for scientific coding and interaction with tools, available in **4B, 9B, and 72B** sizes. Each model is supervised fine-tuned with [ms-swift](https://github.com/modelscope/ms-swift) and released as full BF16 weights with the final LoRA adapter merged, together with its configuration and tokenizer.
25
 
26
  The **training dataset is Codex trajectories**, sourced from [ScienceIDE](https://github.com/aitofound/ScienceIDE).
 
111
  | PEFT | 0.20.0 |
112
  | Datasets | 4.8.4 |
113
  | Tokenizers | 0.23.2 |
114
+ | Accelerate | 1.14.0 |