Skip to content
Navigation Menu
Sign in
Appearance settings
Platform
AI CODE CREATION
GitHub Copilot
Write better code with AI
GitHub Copilot app
Direct agents from issue to merge
MCP Registry
Integrate external tools
DEVELOPER WORKFLOWS
Actions
Automate any workflow
Codespaces
Instant dev environments
Issues
Plan and track work
Code Review
Manage code changes
Code Quality
Enforce quality at merge
APPLICATION SECURITY
GitHub Advanced Security
Find and fix vulnerabilities
Code security
Secure your code as you build
Secret protection
Stop leaks before they start
EXPLORE
Why GitHub
Documentation
Blog
Changelog
Marketplace
View all features
Solutions
BY COMPANY SIZE
Enterprises
Small and medium teams
Startups
Nonprofits
BY USE CASE
App Modernization
DevSecOps
DevOps
CI/CD
View all use cases
BY INDUSTRY
Healthcare
Financial services
Manufacturing
Government
View all industries
View all solutions
Resources
EXPLORE BY TOPIC
AI
Software Development
DevOps
Security
View all topics
EXPLORE BY TYPE
Customer stories
Events & webinars
Ebooks & reports
Business insights
GitHub Skills
SUPPORT & SERVICES
Documentation
Customer support
Community forum
Trust center
Partners
View all resources
Open Source
COMMUNITY
GitHub Sponsors
Fund open source developers
PROGRAMS
Security Lab
Maintainer Community
GitHub Stars
Archive Program
REPOSITORIES
Topics
Trending
Collections
Enterprise
ENTERPRISE SOLUTIONS
Enterprise platform
AI-powered developer platform
AVAILABLE ADD-ONS
GitHub Advanced Security
Enterprise-grade security features
Copilot for Business
Enterprise-grade AI features
Premium Support
Enterprise-grade 24/7 support
Pricing
Search
/
Sign in
Sign up
Appearance settings
You signed in with another tab or window.
Reload
to refresh your session.
You signed out in another tab or window.
Reload
to refresh your session.
You switched accounts on another tab or window.
Reload
to refresh your session.
Dismiss alert
{{ message }}
Niko1221
Strata
Repository navigation
Code
Issues
295
(295)
Pull requests
381
(381)
Actions
Projects
Security and quality
Insights
More
items
All pull requests
Labels
10
(10)
Milestones
0
(0)
New pull request
Search pull requests
is
:
pr
state
:
open
is:pr state:open
Clear filter
Search
Pull requests
Open
381
(381)
Closed
730
(730)
Author
Label
Projects
Milestones
Reviews
Assignee
Sort by
Newest
descending
Sort…
Sort by Newest descending
More items
Comfortable display density
Compact display density
docker image cuda build fix + cuda12 extension
#1848
·
rafal-prasal
opened
Oct 10, 2026
Contributor
·
·
serve: an opt-in VRAM headroom for --expert-cache auto (#1549)
#1841
·
demetree
opened
Oct 10, 2026
Contributor
·
·
serve: a reply that ends inside its thinking with no answer is continued (#1814)
#1840
·
47Hunter47
opened
Oct 10, 2026
·
·
1
1
fix(serve): preserve literal closing thinking tags
#1839
·
CC-David-CC
opened
Oct 10, 2026
Contributor
·
·
Resident RAM mode: the RAM copy follows the conversation (
STRATA_RAM_ADAPT
, opt-in)
#1838
·
SanghunYun
opened
Oct 10, 2026
·
·
1
serve: a Chinese / English language switch for the web app (adapted to v0.1.41)
#1836
·
hogtinbao
opened
Oct 10, 2026
·
·
Community benchmark: 2x RTX PRO 4500 Blackwell, Swift IQ3_XXS, engine 0.1.41 against 0.1.40.1 (single stream to 248k, --batch 8 sweep)
#1834
·
qni-live
opened
Oct 10, 2026
Contributor
·
·
Add optional isolated TeX Live rendering to web chat
#1833
·
CC-David-CC
opened
Oct 10, 2026
Contributor
·
·
Add locally bundled MathJax rendering to web chat
#1832
·
CC-David-CC
opened
Oct 10, 2026
Contributor
·
·
Add locally bundled KaTeX rendering to web chat
#1830
·
CC-David-CC
opened
Oct 10, 2026
Contributor
·
·
Add shared LaTeX/web math core without a rendering engine
#1829
·
CC-David-CC
opened
Oct 10, 2026
Contributor
·
·
sycl: register iq_parity and ple_parity like the top-level build
#1828
·
Momoyeyu
opened
Oct 10, 2026
·
·
1
qsa_select: prompt-path block scores with a thread per query head (below sm_80, same bits)
#1825
·
xrip
opened
Oct 10, 2026
·
·
cpu: Q2_0 rows on AVX-only CPUs (q2_avx1.cpp) instead of ggml-cpu's scalar dot
#1824
·
xrip
opened
Oct 10, 2026
·
·
tokenizer: a long message's ids remembered whole, on top of the piece cache
#1823
·
xrip
opened
Oct 10, 2026
·
·
prefill: gate/up MMQ reads the experts in place (CUDA; patched copy of llama.cpp mmq.cuh)
#1822
·
xrip
opened
Oct 10, 2026
·
·
qsa_prompt_attn: int8 K/V prefetched in registers and sharing one buffer (Turing / sm_75)
#1821
·
xrip
opened
Oct 10, 2026
·
·
qwen35moe follow-ups: VRAM cap + quality mode, MTP for F16 draft tensors, chunked GDN, 256/8 route-resident
#1820
·
atomicmilkshake
opened
Oct 10, 2026
·
·
fix(cpu): make IQ expert dispatch independent of draft grouping
#1818
·
CC-David-CC
opened
Oct 10, 2026
Contributor
·
·
sycl: use the source version for setup model checks
#1817
·
nekomario28
opened
Oct 10, 2026
·
3 tasks done
·
·
1
Opt-in OTLP traces of every /v1 request (0.1.41), plus the docs pages 0.1.41 left uncovered
#1816
·
Buddychek
opened
Oct 10, 2026
·
·
test: gate greedy output across speculative windows and TCP
#1813
·
CC-David-CC
opened
Oct 10, 2026
Contributor
·
·
build: link CUDA MMQ target to the CUDA driver API
#1812
·
CC-David-CC
opened
Oct 10, 2026
Contributor
·
·
Community benchmark: RX 9070 16 GB (gfx1201) on Windows 11, IQ3_XXS, engine 0.1.41
#1811
·
spitzerd
opened
Oct 10, 2026
·
·
feat: persist conversation checkpoints with RAM and disk restore
#1810
·
CC-David-CC
opened
Oct 10, 2026
Contributor
·
·
Previous
1
2
3
4
5
6
7
…
16
Next
You can’t perform that action at this time.