Skip to content
Navigation Menu
Sign in
Appearance settings
Platform
AI CODE CREATION
GitHub Copilot
Write better code with AI
GitHub Copilot app
Direct agents from issue to merge
MCP Registry
Integrate external tools
DEVELOPER WORKFLOWS
Actions
Automate any workflow
Codespaces
Instant dev environments
Issues
Plan and track work
Code Review
Manage code changes
Code Quality
Enforce quality at merge
APPLICATION SECURITY
GitHub Advanced Security
Find and fix vulnerabilities
Code security
Secure your code as you build
Secret protection
Stop leaks before they start
EXPLORE
Why GitHub
Documentation
Blog
Changelog
Marketplace
View all features
Solutions
BY COMPANY SIZE
Enterprises
Small and medium teams
Startups
Nonprofits
BY USE CASE
App Modernization
DevSecOps
DevOps
CI/CD
View all use cases
BY INDUSTRY
Healthcare
Financial services
Manufacturing
Government
View all industries
View all solutions
Resources
EXPLORE BY TOPIC
AI
Software Development
DevOps
Security
View all topics
EXPLORE BY TYPE
Customer stories
Events & webinars
Ebooks & reports
Business insights
GitHub Skills
SUPPORT & SERVICES
Documentation
Customer support
Community forum
Trust center
Partners
View all resources
Open Source
COMMUNITY
GitHub Sponsors
Fund open source developers
PROGRAMS
Security Lab
Maintainer Community
GitHub Stars
Archive Program
REPOSITORIES
Topics
Trending
Collections
Enterprise
ENTERPRISE SOLUTIONS
Enterprise platform
AI-powered developer platform
AVAILABLE ADD-ONS
GitHub Advanced Security
Enterprise-grade security features
Copilot for Business
Enterprise-grade AI features
Premium Support
Enterprise-grade 24/7 support
Pricing
Search
/
Sign in
Sign up
Appearance settings
You signed in with another tab or window.
Reload
to refresh your session.
You signed out in another tab or window.
Reload
to refresh your session.
You switched accounts on another tab or window.
Reload
to refresh your session.
Dismiss alert
{{ message }}
Helsinki-NLP
/
mammoth
Public
Notifications
You must be signed in to change notification settings
Fork
4
Star
36
Code
Issues
50
Pull requests
3
Actions
Projects
Security and quality
0
Insights
Additional navigation options
Code
Issues
Pull requests
Actions
Projects
Security and quality
Insights
Actions: Helsinki-NLP/mammoth
Actions
All workflows
Workflows
Deploy Docs & Publish to PyPi
Deploy Docs & Publish to PyPi
Lint & Tests
Lint & Tests
Show more workflows...
Management
Caches
Deployments
All workflows
All workflows
Actions
Loading...
Loading
Sorry, something went wrong.
Uh oh!
There was an error while loading.
Please reload this page
.
will be ignored since log searching is not yet available
Showing runs from all workflows
will be ignored since log searching is not yet available
560 workflow runs
560 workflow runs
Workflow
Filter by Workflow
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching workflows.
Event
Filter by Event
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching events.
Status
Filter by Status
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching statuses.
Branch
Filter by Branch
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching branches.
Actor
Filter by Actor
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching users.
Replace blanket xavier/kaiming init with transformer-scale embedding …
Lint & Tests
#1138:
Commit
f1a9a5b
pushed by
chaowang0524
20s
pytorch_backend
pytorch_backend
20s
View workflow file
Fix ff_mult not auto-defaulting to 4.0 for gelu activation
Lint & Tests
#1137:
Commit
6cb831c
pushed by
chaowang0524
23s
pytorch_backend
pytorch_backend
23s
View workflow file
Add xavier/kaiming weight init for native transformer backend
Lint & Tests
#1136:
Commit
e03f340
pushed by
chaowang0524
23s
pytorch_backend
pytorch_backend
23s
View workflow file
Replace blanket xavier/kaiming init with transformer-scale embedding …
Lint & Tests
#1135:
Commit
f91fd6b
pushed by
chaowang0524
21s
embedding-transformer-scale-init
embedding-transformer-scale-init
21s
View workflow file
Fix ff_mult not auto-defaulting to 4.0 for gelu activation
Lint & Tests
#1134:
Commit
6cb831c
pushed by
chaowang0524
23s
pytorch_backend
pytorch_backend
23s
View workflow file
Fix crash when metric_for_best_model is unavailable during validation
Lint & Tests
#1133:
Commit
6490840
pushed by
chaowang0524
24s
cherry-pick/pytorch-backend-fixes
cherry-pick/pytorch-backend-fixes
24s
View workflow file
Fix crash when metric_for_best_model is unavailable during validation
Lint & Tests
#1132:
Commit
23cc85e
pushed by
chaowang0524
26s
pytorch_backend
pytorch_backend
26s
View workflow file
Fix ModelSaver reading save_strategy from stale checkpoint opts
Lint & Tests
#1131:
Commit
bb7eed4
pushed by
chaowang0524
23s
pytorch_backend
pytorch_backend
23s
View workflow file
Fix checkpoint metadata write failing on directory-style save_model p…
Lint & Tests
#1130:
Commit
fabef30
pushed by
chaowang0524
20s
pytorch_backend
pytorch_backend
20s
View workflow file
Save converted tokenizer vocab alongside Gemma3 conversion checkpoint
Lint & Tests
#1129:
Commit
0056717
pushed by
chaowang0524
24s
pytorch_backend
pytorch_backend
24s
View workflow file
Freeze more of the decoder body when using freeze_decoder*
Lint & Tests
#1128:
Commit
75df9e6
pushed by
chaowang0524
21s
pytorch_backend
pytorch_backend
21s
View workflow file
Updated the gemma3 conversion doc
Lint & Tests
#1127:
Commit
ac86c77
pushed by
chaowang0524
25s
pytorch_backend
pytorch_backend
25s
View workflow file
Updated "torch.dtype" to "dtype" to fit the transformers library on LUMI
Lint & Tests
#1126:
Commit
4e84a47
pushed by
chaowang0524
22s
pytorch_backend
pytorch_backend
22s
View workflow file
Updated "torch.dtype" to "dtype" to fit the transformers library on LUMI
Lint & Tests
#1125:
Commit
4e84a47
pushed by
chaowang0524
26s
gemma3-native-decoder-only
gemma3-native-decoder-only
26s
View workflow file
Updated model dtype from fp32 to bf16
Lint & Tests
#1124:
Commit
2797a28
pushed by
chaowang0524
19s
gemma3-native-decoder-only
gemma3-native-decoder-only
19s
View workflow file
Add Gemma3-270M -> Mammoth native-backend conversion (fake-encoder an…
Lint & Tests
#1123:
Commit
d4634bc
pushed by
chaowang0524
26s
gemma3-native-decoder-only
gemma3-native-decoder-only
26s
View workflow file
Merge fix/model-saver-collective-desync into pytorch_backend
Lint & Tests
#1122:
Commit
33f2f7c
pushed by
chaowang0524
20s
pytorch_backend
pytorch_backend
20s
View workflow file
Fix distributed collective desync in model_saver metric extraction
Lint & Tests
#1121:
Commit
9d60951
pushed by
chaowang0524
24s
fix/model-saver-collective-desync
fix/model-saver-collective-desync
24s
View workflow file
Validate weight=0 tasks independently and log their own samples
Lint & Tests
#1120:
Commit
ecde0ee
pushed by
chaowang0524
24s
pytorch_backend
pytorch_backend
24s
View workflow file
Validate weight=0 tasks independently and log their own samples
Lint & Tests
#1119:
Commit
ecde0ee
pushed by
chaowang0524
23s
fix/per-task-validation
fix/per-task-validation
23s
View workflow file
Updated the 'metric_for_best_model' to be BLEU
Lint & Tests
#1118:
Commit
732e325
pushed by
chaowang0524
21s
pytorch_backend
pytorch_backend
21s
View workflow file
Merge branch 'fix/early-stopping' into dev
Lint & Tests
#1117:
Commit
176fe62
pushed by
chaowang0524
26s
dev
dev
26s
View workflow file
reshaped eval2
Lint & Tests
#1116:
Commit
0e7e258
pushed by
amikael
27s
feat/task-eval
feat/task-eval
27s
View workflow file
updated rudimentary documentation
Lint & Tests
#1115:
Commit
b237f3b
pushed by
amikael
28s
feat/task-eval
feat/task-eval
28s
View workflow file
added rudimentary documentation
Lint & Tests
#1114:
Commit
37ccfbc
pushed by
amikael
24s
feat/task-eval
feat/task-eval
24s
View workflow file
Previous
1
2
3
4
5
…
22
23
Next
You can’t perform that action at this time.