Skip to content
Navigation Menu
Sign in
Appearance settings
Platform
AI CODE CREATION
GitHub Copilot
Write better code with AI
GitHub Copilot app
Direct agents from issue to merge
MCP Registry
Integrate external tools
DEVELOPER WORKFLOWS
Actions
Automate any workflow
Codespaces
Instant dev environments
Issues
Plan and track work
Code Review
Manage code changes
Code Quality
Enforce quality at merge
APPLICATION SECURITY
GitHub Advanced Security
Find and fix vulnerabilities
Code security
Secure your code as you build
Secret protection
Stop leaks before they start
EXPLORE
Why GitHub
Documentation
Blog
Changelog
Marketplace
View all features
Solutions
BY COMPANY SIZE
Enterprises
Small and medium teams
Startups
Nonprofits
BY USE CASE
App Modernization
DevSecOps
DevOps
CI/CD
View all use cases
BY INDUSTRY
Healthcare
Financial services
Manufacturing
Government
View all industries
View all solutions
Resources
EXPLORE BY TOPIC
AI
Software Development
DevOps
Security
View all topics
EXPLORE BY TYPE
Customer stories
Events & webinars
Ebooks & reports
Business insights
GitHub Skills
SUPPORT & SERVICES
Documentation
Customer support
Community forum
Trust center
Partners
View all resources
Open Source
COMMUNITY
GitHub Sponsors
Fund open source developers
PROGRAMS
Security Lab
Maintainer Community
Accelerator
GitHub Stars
Archive Program
REPOSITORIES
Topics
Trending
Collections
Enterprise
ENTERPRISE SOLUTIONS
Enterprise platform
AI-powered developer platform
AVAILABLE ADD-ONS
GitHub Advanced Security
Enterprise-grade security features
Copilot for Business
Enterprise-grade AI features
Premium Support
Enterprise-grade 24/7 support
Pricing
Search
/
Sign in
Sign up
Appearance settings
You signed in with another tab or window.
Reload
to refresh your session.
You signed out in another tab or window.
Reload
to refresh your session.
You switched accounts on another tab or window.
Reload
to refresh your session.
Dismiss alert
{{ message }}
Uh oh!
There was an error while loading.
Please reload this page
.
open-infra-ai
/
tiny-llm
Public
Notifications
You must be signed in to change notification settings
Fork
0
Star
0
Code
Issues
0
Pull requests
0
Actions
Projects
Security and quality
0
Insights
Additional navigation options
Code
Issues
Pull requests
Actions
Projects
Security and quality
Insights
Actions: open-infra-ai/tiny-llm
Actions
All workflows
Workflows
CI
CI
GitHub Pages
GitHub Pages
Release
Release
Show more workflows...
Management
Caches
Deployments
All workflows
All workflows
Actions
Loading...
Loading
Sorry, something went wrong.
Uh oh!
There was an error while loading.
Please reload this page
.
will be ignored since log searching is not yet available
Showing runs from all workflows
will be ignored since log searching is not yet available
121 workflow runs
121 workflow runs
Workflow
Filter by Workflow
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching workflows.
Event
Filter by Event
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching events.
Status
Filter by Status
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching statuses.
Branch
Filter by Branch
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching branches.
Actor
Filter by Actor
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching users.
test(ci): skip GPU-only cases without a device
CI
#87:
Commit
00cc604
pushed by
holtwood
9m 9s
master
master
9m 9s
View workflow file
ci: fix CUDA setup under Node.js 24
CI
#86:
Commit
b6e143b
pushed by
holtwood
9m 56s
master
master
9m 56s
View workflow file
style: align sources with clang-format 18
CI
#85:
Commit
606e493
pushed by
holtwood
25s
master
master
25s
View workflow file
docs(performance): publish clean CUDA Graph A/B evidence
GitHub Pages
#32:
Commit
26b74fe
pushed by
holtwood
1m 5s
master
master
1m 5s
View workflow file
docs(performance): publish clean CUDA Graph A/B evidence
CI
#84:
Commit
26b74fe
pushed by
holtwood
22s
master
master
22s
View workflow file
docs(kv-cache): document paged KV strategy 1
GitHub Pages
#31:
Commit
22a4cae
pushed by
holtwood
56s
master
master
56s
View workflow file
chore(build/misc): cuda arch 70, sizeof semantics, drop redundant cast
CI
#83:
Commit
0a293a0
pushed by
holtwood
26s
master
master
26s
View workflow file
docs: normalize org links to open-infra-ai
GitHub Pages
#30:
Commit
b8af256
pushed by
holtwood
49s
master
master
49s
View workflow file
docs(kv-cache): document getUsedMemory slot-based accounting
CI
#82:
Commit
0f711d6
pushed by
holtwood
17s
master
master
17s
View workflow file
fix(runtime): free GPU buffers when LayerWorkspace::allocate throws
CI
#81:
Commit
1e58f30
pushed by
holtwood
17s
master
master
17s
View workflow file
fix(runtime): release GPU resources on construction exception
CI
#80:
Commit
72d8264
pushed by
holtwood
16s
master
master
16s
View workflow file
docs: use aicl-lab GitHub org in public links
GitHub Pages
#29:
Commit
70d034e
pushed by
holtwood
1m 3s
master
master
1m 3s
View workflow file
docs: use aicl-lab GitHub org in public links
CI
#79:
Commit
70d034e
pushed by
holtwood
23s
master
master
23s
View workflow file
docs: mark paged KV (strategy 1) enabled in README, ffi.h and develop…
CI
#78:
Commit
1c0abdd
pushed by
holtwood
18s
master
master
18s
View workflow file
docs(perf): decode optimization report and updated benchmark snapshot
GitHub Pages
#28:
Commit
6d0471e
pushed by
holtwood
58s
master
master
58s
View workflow file
docs(perf): decode optimization report and updated benchmark snapshot
CI
#77:
Commit
6d0471e
pushed by
holtwood
19s
master
master
19s
View workflow file
refactor(runtime): extract finalNormAndComputeLogits shared by engine…
CI
#76:
Commit
3ddafcc
pushed by
holtwood
22s
master
master
22s
View workflow file
refactor(runtime): extract finalNormAndComputeLogits shared by engine…
GitHub Pages
#27:
Commit
3ddafcc
pushed by
holtwood
1m 4s
master
master
1m 4s
View workflow file
feat: KV cache 显式 seq_id + C ABI 契约完善(接入配套)
CI
#75:
Commit
058dc8b
pushed by
holtwood
23s
master
master
23s
View workflow file
feat: 导出 C ABI 执行后端(里程碑 2)
CI
#74:
Commit
0898e7b
pushed by
holtwood
14s
master
master
14s
View workflow file
feat: stage-1 GPU end-to-end generation on Qwen2.5-0.5B
CI
#73:
Commit
92b6699
pushed by
holtwood
17s
master
master
17s
View workflow file
docs: mark tokenizer complete in roadmap, readme, and changelog
CI
#72:
Commit
94c8d12
pushed by
holtwood
16s
master
master
16s
View workflow file
docs: record real-model verification results and stage-1 progress
CI
#71:
Commit
5b807ba
pushed by
holtwood
9m 11s
master
master
9m 11s
View workflow file
fix: use absolute GitHub URLs for ROADMAP links in docs site
GitHub Pages
#26:
Commit
b0c5ec7
pushed by
holtwood
11m 5s
master
master
11m 5s
View workflow file
docs: add ROADMAP with phased interview-oriented plan
GitHub Pages
#25:
Commit
68e9d32
pushed by
holtwood
33s
master
master
33s
View workflow file
Previous
1
2
3
4
5
Next
You can’t perform that action at this time.