Skip to content
Navigation Menu
Sign in
Appearance settings
Platform
AI CODE CREATION
GitHub Copilot
Write better code with AI
GitHub Copilot app
Direct agents from issue to merge
MCP Registry
Integrate external tools
DEVELOPER WORKFLOWS
Actions
Automate any workflow
Codespaces
Instant dev environments
Issues
Plan and track work
Code Review
Manage code changes
Code Quality
Enforce quality at merge
APPLICATION SECURITY
GitHub Advanced Security
Find and fix vulnerabilities
Code security
Secure your code as you build
Secret protection
Stop leaks before they start
EXPLORE
Why GitHub
Documentation
Blog
Changelog
Marketplace
View all features
Solutions
BY COMPANY SIZE
Enterprises
Small and medium teams
Startups
Nonprofits
BY USE CASE
App Modernization
DevSecOps
DevOps
CI/CD
View all use cases
BY INDUSTRY
Healthcare
Financial services
Manufacturing
Government
View all industries
View all solutions
Resources
EXPLORE BY TOPIC
AI
Software Development
DevOps
Security
View all topics
EXPLORE BY TYPE
Customer stories
Events & webinars
Ebooks & reports
Business insights
GitHub Skills
SUPPORT & SERVICES
Documentation
Customer support
Community forum
Trust center
Partners
View all resources
Open Source
COMMUNITY
GitHub Sponsors
Fund open source developers
PROGRAMS
Security Lab
Maintainer Community
GitHub Stars
Archive Program
REPOSITORIES
Topics
Trending
Collections
Enterprise
ENTERPRISE SOLUTIONS
Enterprise platform
AI-powered developer platform
AVAILABLE ADD-ONS
GitHub Advanced Security
Enterprise-grade security features
Copilot for Business
Enterprise-grade AI features
Premium Support
Enterprise-grade 24/7 support
Pricing
Search
/
Sign in
Sign up
Appearance settings
You signed in with another tab or window.
Reload
to refresh your session.
You signed out in another tab or window.
Reload
to refresh your session.
You switched accounts on another tab or window.
Reload
to refresh your session.
Dismiss alert
{{ message }}
FlashML-org
/
FreeToken
Public
Notifications
You must be signed in to change notification settings
Fork
1.1k
Star
11.9k
Code
Issues
164
Pull requests
133
Actions
Projects
Security and quality
0
Insights
Additional navigation options
Code
Issues
Pull requests
Actions
Projects
Security and quality
Insights
Actions: FlashML-org/FreeToken
Actions
All workflows
Workflows
Build container image
Build container image
Copilot
Copilot
Label issues
Label issues
Nightly wheels
Nightly wheels
Release wheels
Release wheels
Unit tests (NVIDIA)
Unit tests (NVIDIA)
Show more workflows...
Management
Caches
Deployments
All workflows
All workflows
Actions
Loading...
Loading
Sorry, something went wrong.
Uh oh!
There was an error while loading.
Please reload this page
.
will be ignored since log searching is not yet available
Showing runs from all workflows
will be ignored since log searching is not yet available
54 workflow runs
54 workflow runs
Workflow
Filter by Workflow
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching workflows.
Event
Filter by Event
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching events.
Status
Filter by Status
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching statuses.
Branch
Filter by Branch
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching branches.
Actor
Filter by Actor
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching users.
Nightly wheels
Nightly wheels
#30:
Scheduled
7s
main
main
7s
View workflow file
[Bug] Checkpoint conversion fails for Qwen3.6-35B FP8: Tensor size mismatch in dequant_fp8_weight
Label issues
#14:
Issue
#402
opened by
quanrennsxsb
7s
7s
View workflow file
--moe-cache-auto leaves no room for prefill: boots and serves decode, then the first 8k prefill OOMs and kills the worker (NVFP4 dense)
Label issues
#13:
Issue
#401
opened by
salekseev
7s
7s
View workflow file
Help/pr 132
Unit tests (NVIDIA)
#2:
Pull request
#400
opened by
samuelishida
Action required
samuelishida:help/pr-132
samuelishida:help/pr-132
Action required
View #400
View workflow file
Running Copilot Code Review
Copilot
#5:
by
Copilot
AI
4m 1s
main
main
4m 1s
5090+256G内存无法运行红帽的GLM5.3FlashNVFP4
Label issues
#12:
Issue
#397
opened by
Zanry
43s
43s
View workflow file
[Feature] int4 routed experts: AWQ -> Q4_1 and compressed-tensors pack-quantized -> Q4_0, with no new CUDA kernel (offering a PR)
Label issues
#11:
Issue
#396
opened by
salekseev
9s
9s
View workflow file
[Bug] --max-output-tokens is ignored by /v1/chat/completions and /v1/messages (only /v1/responses honours it)
Label issues
#10:
Issue
#395
opened by
salekseev
8s
8s
View workflow file
Support official nvidia/Qwen3.8-Flash-Next-NVFP4 model
Label issues
#9:
Issue
#394
opened by
lonestriker
7s
7s
View workflow file
Nightly wheels
Nightly wheels
#29:
Scheduled
8s
main
main
8s
View workflow file
FreeToken-Intel: 35B MoE, live tokens off an Arc Pro B70 (SYCL / XPU port)
Label issues
#8:
Issue
#299
edited by
aangelinsf
8s
8s
View workflow file
Desktop engine installer ignores existing FreeToken engine path and always installs a new venv in AppData
Label issues
#7:
Issue
#387
opened by
MSM-App
9s
9s
View workflow file
Running Copilot Code Review
Copilot
#4:
by
Copilot
AI
2m 40s
main
main
2m 40s
Running Copilot Code Review
Copilot
#3:
by
Copilot
AI
5m 5s
main
main
5m 5s
Build on Windows?
Label issues
#6:
Issue
#384
opened by
XeonG
9s
9s
View workflow file
--moe-cache-auto ignores --num-tokens / --num-pages: expert fill overruns the budget and the KV pool OOMs at boot
Label issues
#5:
Issue
#383
opened by
gdevenyi
7s
7s
View workflow file
Move the The Cache config element in the console window for QOL improvement.
Label issues
#4:
Issue
#382
opened by
bmgjet
9s
9s
View workflow file
Nightly wheels
Nightly wheels
#28:
Scheduled
3m 47s
main
main
3m 47s
View workflow file
[Bug]: qwen3_5_moe fails to convert/load compressed-tensors NVFP4 checkpoints with model.language_model prefix
Label issues
#3:
Issue
#381
opened by
silvertakana
8s
8s
View workflow file
ci(container): настроить сборку образа
Build container image
#1:
Pull request
#380
opened by
Fgeeha
Action required
Fgeeha:ci/docker-images
Fgeeha:ci/docker-images
Action required
View #380
View workflow file
ci(container): publish official Docker images to GHCR
Label issues
#2:
Issue
#379
opened by
Fgeeha
7s
7s
View workflow file
could not encode request: Unexpected message role.
Label issues
#1:
Issue
#376
opened by
MCJEModder2026
7s
7s
View workflow file
Nightly wheels
Nightly wheels
#27:
Scheduled
3m 43s
main
main
3m 43s
View workflow file
Nightly wheels
Nightly wheels
#26:
Scheduled
3m 14s
main
main
3m 14s
View workflow file
feat(rocm): AMD ROCm runtime for RDNA3/4 with native Qwen3.5-MoE GGUF decode parity
Unit tests (NVIDIA)
#1:
Pull request
#217
synchronize by
samuelishida
Action required
samuelishida:feat/amd-rocm-gfx1100-support
samuelishida:feat/amd-rocm-gfx1100-support
Action required
View #217
View workflow file
Previous
1
2
3
Next
You can’t perform that action at this time.