Skip to content
Navigation Menu
Sign in
Appearance settings
Platform
AI CODE CREATION
GitHub Copilot
Write better code with AI
GitHub Copilot app
Direct agents from issue to merge
MCP Registry
Integrate external tools
DEVELOPER WORKFLOWS
Actions
Automate any workflow
Codespaces
Instant dev environments
Issues
Plan and track work
Code Review
Manage code changes
Code Quality
Enforce quality at merge
APPLICATION SECURITY
GitHub Advanced Security
Find and fix vulnerabilities
Code security
Secure your code as you build
Secret protection
Stop leaks before they start
EXPLORE
Why GitHub
Documentation
Blog
Changelog
Marketplace
View all features
Solutions
BY COMPANY SIZE
Enterprises
Small and medium teams
Startups
Nonprofits
BY USE CASE
App Modernization
DevSecOps
DevOps
CI/CD
View all use cases
BY INDUSTRY
Healthcare
Financial services
Manufacturing
Government
View all industries
View all solutions
Resources
EXPLORE BY TOPIC
AI
Software Development
DevOps
Security
View all topics
EXPLORE BY TYPE
Customer stories
Events & webinars
Ebooks & reports
Business insights
GitHub Skills
SUPPORT & SERVICES
Documentation
Customer support
Community forum
Trust center
Partners
View all resources
Open Source
COMMUNITY
GitHub Sponsors
Fund open source developers
PROGRAMS
Security Lab
Maintainer Community
Accelerator
GitHub Stars
Archive Program
REPOSITORIES
Topics
Trending
Collections
Enterprise
ENTERPRISE SOLUTIONS
Enterprise platform
AI-powered developer platform
AVAILABLE ADD-ONS
GitHub Advanced Security
Enterprise-grade security features
Copilot for Business
Enterprise-grade AI features
Premium Support
Enterprise-grade 24/7 support
Pricing
Search
/
Sign in
Sign up
Appearance settings
You signed in with another tab or window.
Reload
to refresh your session.
You signed out in another tab or window.
Reload
to refresh your session.
You switched accounts on another tab or window.
Reload
to refresh your session.
Dismiss alert
{{ message }}
FluffyAIcode
/
Kakeya-LLM-Inference-engine
Public
Notifications
You must be signed in to change notification settings
Fork
0
Star
0
Code
Issues
0
Pull requests
19
Actions
Projects
Security and quality
0
Insights
Additional navigation options
Code
Issues
Pull requests
Actions
Projects
Security and quality
Insights
Actions: FluffyAIcode/Kakeya-LLM-Inference-engine
Actions
All workflows
Workflows
Auto-label needs-mac-m4
Auto-label needs-mac-m4
CI
CI
Integration (Mac M4)
Integration (Mac M4)
Mac bridge
Mac bridge
Show more workflows...
Management
Caches
Integration (Mac M4)
Integration (Mac M4)
Actions
Loading...
Loading
Sorry, something went wrong.
Uh oh!
There was an error while loading.
Please reload this page
.
will be ignored since log searching is not yet available
Show workflow options
Create status badge
Create status badge
Loading
Uh oh!
There was an error while loading.
Please reload this page
.
integration.yaml
will be ignored since log searching is not yet available
274 workflow runs
274 workflow runs
Event
Filter by Event
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching events.
Status
Filter by Status
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching statuses.
Branch
Filter by Branch
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching branches.
Actor
Filter by Actor
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching users.
Harden Primary decode lifecycle and isolate MLX worker
Integration (Mac M4)
#224:
Pull request
#227
opened by
FluffyAIcode
1m 52s
agent/primary-decode-worker-acceptance-integration-0721
agent/primary-decode-worker-acceptance-integration-0721
1m 52s
View #227
View workflow file
fix(prefill): inject immutable retained-token cap
Integration (Mac M4)
#223:
Pull request
#201
synchronize by
FluffyAIcode
1m 48s
AgentMemory/explicit-retained-token-cap-0719
AgentMemory/explicit-retained-token-cap-0719
1m 48s
View #201
View workflow file
fix(prefill): inject immutable retained-token cap
Integration (Mac M4)
#222:
Pull request
#201
opened by
FluffyAIcode
1m 57s
AgentMemory/explicit-retained-token-cap-0719
AgentMemory/explicit-retained-token-cap-0719
1m 57s
View #201
View workflow file
fix(prefill): cap reservation at sliding window
Integration (Mac M4)
#221:
Pull request
#198
opened by
FluffyAIcode
2m 4s
AgentMemory/window-aware-snapshot-estimate-0718
AgentMemory/window-aware-snapshot-estimate-0718
2m 4s
View #198
View workflow file
perf(prefill): segment long jobs for autoresearch
Integration (Mac M4)
#220:
Pull request
#192
synchronize by
FluffyAIcode
1m 46s
AgentMemory/prefill-autoresearch-segments-0717
AgentMemory/prefill-autoresearch-segments-0717
1m 46s
View #192
View workflow file
perf(prefill): segment long jobs for autoresearch
Integration (Mac M4)
#219:
Pull request
#192
opened by
FluffyAIcode
1m 57s
AgentMemory/prefill-autoresearch-segments-0717
AgentMemory/prefill-autoresearch-segments-0717
1m 57s
View #192
View workflow file
perf(prefill): export only final MLX snapshot
Integration (Mac M4)
#218:
Pull request
#191
opened by
FluffyAIcode
2m 1s
AgentMemory/final-only-prefill-snapshot-0717
AgentMemory/final-only-prefill-snapshot-0717
2m 1s
View #191
View workflow file
fix(prefill): preserve unrelated snapshots during jobs
Integration (Mac M4)
#217:
Pull request
#188
opened by
FluffyAIcode
1m 55s
AgentMemory/precise-snapshot-reservation-0717
AgentMemory/precise-snapshot-reservation-0717
1m 55s
View #188
View workflow file
fix(prefill): atomically reserve and publish snapshots
Integration (Mac M4)
#216:
Pull request
#187
synchronize by
FluffyAIcode
1m 49s
AgentMemory/atomic-prefill-snapshot-0717
AgentMemory/atomic-prefill-snapshot-0717
1m 49s
View #187
View workflow file
fix(prefill): atomically reserve and publish snapshots
Integration (Mac M4)
#215:
Pull request
#187
opened by
FluffyAIcode
1m 56s
AgentMemory/atomic-prefill-snapshot-0717
AgentMemory/atomic-prefill-snapshot-0717
1m 56s
View #187
View workflow file
fix(agents): require full-context semantic Critic review
Integration (Mac M4)
#214:
Pull request
#184
opened by
FluffyAIcode
1m 58s
AgentMemory/full-context-critic-0717
AgentMemory/full-context-critic-0717
1m 58s
View #184
View workflow file
fix(agents): bound Critic evidence prefill context
Integration (Mac M4)
#213:
Pull request
#181
opened by
FluffyAIcode
1m 53s
AgentMemory/critic-evidence-window-0717
AgentMemory/critic-evidence-window-0717
1m 53s
View #181
View workflow file
fix(agents): complete responses to EOS and ground critique
Integration (Mac M4)
#212:
Pull request
#178
opened by
FluffyAIcode
2m 5s
AgentMemory/agent-completion-rubric-0716
AgentMemory/agent-completion-rubric-0716
2m 5s
View #178
View workflow file
feat(agents): add Generator-Critic inference demo
Integration (Mac M4)
#211:
Pull request
#175
opened by
FluffyAIcode
1m 51s
AgentMemory/agent-gan-inference-demo-0716
AgentMemory/agent-gan-inference-demo-0716
1m 51s
View #175
View workflow file
feat(bench): add three-phase prefill fleet benchmark
Integration (Mac M4)
#210:
Pull request
#174
opened by
FluffyAIcode
1m 49s
AgentMemory/prefill-benchmark-task-0716
AgentMemory/prefill-benchmark-task-0716
1m 49s
View #174
View workflow file
fix(prefill): resolve promoted final snapshots directly
Integration (Mac M4)
#209:
Pull request
#173
opened by
FluffyAIcode
2m 5s
AgentMemory/sparse-hot-kv-lookup-0716
AgentMemory/sparse-hot-kv-lookup-0716
2m 5s
View #173
View workflow file
fix(prefill): isolate cache budget policy from MLX runtime
Integration (Mac M4)
#208:
Pull request
#172
opened by
FluffyAIcode
1m 47s
AgentMemory/tiered-kv-offload-0716
AgentMemory/tiered-kv-offload-0716
1m 47s
View #172
View workflow file
feat(prefill): add tiered KV hot promotion and offload
Integration (Mac M4)
#207:
Pull request
#171
opened by
FluffyAIcode
1m 55s
AgentMemory/tiered-kv-offload-0716
AgentMemory/tiered-kv-offload-0716
1m 55s
View #171
View workflow file
feat(prefill): enforce decode-only primary mode
Integration (Mac M4)
#206:
Pull request
#170
opened by
FluffyAIcode
1m 59s
AgentMemory/primary-decode-allens-prefill-0716
AgentMemory/primary-decode-allens-prefill-0716
1m 59s
View #170
View workflow file
feat(prefill): add maintenance cache saturation harness
Integration (Mac M4)
#205:
Pull request
#169
opened by
FluffyAIcode
1m 41s
AgentMemory/kv-cache-saturation-harness-0712
AgentMemory/kv-cache-saturation-harness-0712
1m 41s
View #169
View workflow file
feat(prefill): add bit-packed KakeyaLattice snapshots
Integration (Mac M4)
#204:
Pull request
#167
opened by
FluffyAIcode
1m 43s
AgentMemory/prefill-kakeyalattice-codec-0712
AgentMemory/prefill-kakeyalattice-codec-0712
1m 43s
View #167
View workflow file
fix(prefill): keep MLX worker compute thread-affine
Integration (Mac M4)
#203:
Pull request
#165
opened by
FluffyAIcode
1m 40s
AgentMemory/mlx-prefill-worker-thread-affinity-0712
AgentMemory/mlx-prefill-worker-thread-affinity-0712
1m 40s
View #165
View workflow file
feat(prefill): operationalize remote MLX workers
Integration (Mac M4)
#202:
Pull request
#164
opened by
FluffyAIcode
1m 41s
AgentMemory/prefill-worker-production-completion-0712
AgentMemory/prefill-worker-production-completion-0712
1m 41s
View #164
View workflow file
feat(distributed): orchestrate same-model prefill workers and KV memory tier
Integration (Mac M4)
#201:
Pull request
#160
synchronize by
cursor
Bot
2m 18s
AgentMemory/prefill-worker-orchestration-2815
AgentMemory/prefill-worker-orchestration-2815
2m 18s
View #160
View workflow file
feat(distributed): orchestrate same-model prefill workers and KV memory tier
Integration (Mac M4)
#200:
Pull request
#160
synchronize by
cursor
Bot
43s
AgentMemory/prefill-worker-orchestration-2815
AgentMemory/prefill-worker-orchestration-2815
43s
View #160
View workflow file
Previous
1
2
3
4
5
…
10
11
Next
You can’t perform that action at this time.