bio2C v1
Benchmark
A biomedical benchmark containing 2,500 manually curated natural language questions and corresponding Cypher queries over miRNA-KG.
Statistics
| Property |
Value |
| Version |
v1 |
| Domains |
Biomedical |
| Languages |
English |
| Curation |
Manually curated |
| Total NL–Cypher pairs |
2,500 |
| Graph-executable pairs |
2,500 |
| Complexity levels |
5 |
| Complexity definition |
The dataset is divided into five categories: node-level, 1-hop, 2-hop, 3-hop, and advanced. Each category contains 500 pairs. |
Database: miRNA-KG T2C
Endpoint
Best reported results
| Metric |
Score |
Model |
Technique(s) |
| Jaro–Winkler |
94.7% |
GPT-4oFT |
RAG, RAG+O, S+RAG |
| Jaccard |
86.4% |
GPT-4oFT |
RAG |
| Coverage |
84.3% |
GPT-4oFT |
RAG |
| Pass@1 |
89.3% |
GPT-4oFT |
RAG, S+RAG+O |
Reference and license
Notes