Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Ryan Marten
ryanmarten
54
23
70
Follow
Notqty's profile picture
dujun's profile picture
RODRIGUE-ZMICHAEL's profile picture
46 followers
·
99 following
https://ryanmarten.com
ryanmart3n
ryanmarten
ryan-marten
AI & ML interests
None yet
Recent Activity
updated
a dataset
8 days ago
harborframework/terminal-bench-lfs
published
a dataset
8 days ago
harborframework/terminal-bench-lfs
updated
a dataset
12 days ago
harborframework/terminal-bench-3.0
View all activity
Organizations
ryanmarten
's activity
All
Models
Datasets
Spaces
Buckets
Papers
Collections
Community
Posts
Upvotes
Likes
Articles
updated
a dataset
8 days ago
harborframework/terminal-bench-lfs
Preview
•
Updated
8 days ago
•
757
published
a dataset
8 days ago
harborframework/terminal-bench-lfs
Preview
•
Updated
8 days ago
•
757
updated
a dataset
12 days ago
harborframework/terminal-bench-3.0
Benchmark
•
Updated
12 days ago
•
13.3k
•
3
upvoted
a
paper
2 months ago
OpenThoughts-Agent: Data Recipes for Agentic Models
Paper
•
2606.24855
•
Published
Jun 23
•
48
New activity in
harborframework/parity-experiments
4 months ago
Add parity experiments for harveyai/lab
#250 opened 4 months ago by
ryanmarten
liked
a dataset
4 months ago
open-thoughts/AgentTrove
Viewer
•
Updated
May 7
•
1.7M
•
5.1k
•
195
updated
a dataset
4 months ago
harborframework/terminal-bench-2.0
Benchmark
•
Updated
Apr 24
•
37.8k
•
48
New activity in
harborframework/parity-experiments
6 months ago
SpreadsheetBench adapter parity (claude-code + Haiku 4.5, 400 tasks × 3 trials)
2
#106 opened 6 months ago by
ryanmarten
New activity in
harborframework/terminal-bench-2.0
7 months ago
Define 'harbor' as eval framework 🎉
#3 opened 7 months ago by
burtenshaw
Add an eval yaml to integrate this benchmark into Community Evals.
#1 opened 7 months ago by
burtenshaw
published
a dataset
7 months ago
harborframework/terminal-bench-2.0
Benchmark
•
Updated
Apr 24
•
37.8k
•
48
liked
a dataset
7 months ago
zai-org/terminal-bench-2-verified
Updated
9 days ago
•
16.7k
•
82
liked
a dataset
9 months ago
open-thoughts/OpenThoughts-Agent-v1-SFT
Viewer
•
Updated
Jan 27
•
15.2k
•
2.85k
•
104
updated
a Space
9 months ago
Running
README
🦀
liked
a dataset
10 months ago
jupyter-agent/jupyter-agent-dataset
Viewer
•
Updated
Sep 10, 2025
•
95.8k
•
1.52k
•
172
updated
2 datasets
about 1 year ago
ryanmarten/OpenThoughts-1k-sample
Viewer
•
Updated
Aug 31, 2025
•
2k
•
1.34M
•
50
open-thoughts/OpenThoughts-114k
Viewer
•
Updated
Aug 31, 2025
•
228k
•
112k
•
908
published
a dataset
about 1 year ago
ryanmarten/OpenThoughts-1k-sample
Viewer
•
Updated
Aug 31, 2025
•
2k
•
1.34M
•
50
liked
a dataset
about 1 year ago
SWE-bench/SWE-smith-trajectories
Viewer
•
Updated
Jul 19, 2025
•
76k
•
9.32k
•
78
liked
a Space
about 1 year ago
Running
6
OpenThoughts Benchmark Explorer
📊
6
Explore benchmark correlations and model performance
Load more