Skip to content
View rmems's full-sized avatar

Highlights

  • Pro

Organizations

@Limen-Neural

Block or report rmems

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
.github/profile/README.md

Raul Montoya Cardenas

San Marcos, Texas · montoyaraul34@gmail.com


Making large models run on hardware you can actually own.

My core research is SAAQ (Spiking Adaptive Activity Quantization) — compressing large-scale models so they run on consumer GPUs by turning them into spiking neural network representations. I also work on model interpretability, quantization benchmarking, and a modular neuromorphic stack under Limen Neural.


🧠 Active Projects

Project Description Stack
corinth-canal SAAQ reference loop: telemetry → spiking layer → projector/router → latent calibration Rust · CUDA
xai-dissect Static structural analysis of Grok-family open-weight MoE checkpoints Rust
grok-ozempic Ternary SNN-inspired quantization for Grok-scale MoE routing fidelity Rust
magere-brug SAAQ experiment lab: manifests, recipes, artifact registry Rust
Surrogate_Viz.jl Symbolic regression + validation dashboards for SAAQ telemetry Julia
XAIDissect_Viz.jl Interactive Grok-1 MoE atmosphere from xai-dissect reports Julia · Makie
myelin-accelerator Blackwell-first CUDA kernels for neuromorphic inference Rust · CUDA
agoge-forger Local-first GPU training forge (QLoRA/LoRA, RTX 5080-aware) Python · PyTorch
combine-for-AI Neutral benchmark harness for model quantization experiments Python
Dioscuri-Cloud Cloud ML lab: smoke tests, cost ledger, multi-provider runbooks Terraform · Ops
gaming-telemetry High-frequency GPU/CPU telemetry → Parquet for SNN training Rust · NVML
spike-viz Visualize SNN encodings from axon-encoder exports Python · PyTorch
NeuralForge-Memory Personal RAG + vector DB + MCP memory (Hermes tutor) Rust · MCP

🧬 Limen Neural

Limen Neural is my neuromorphic library org — hard repo boundaries, modular crates you can reuse, dual MIT/Apache-2.0 where applicable.


🤖 Multi-Agent Engineering

I ship research with a multi-agent engineering stack across Grok Build, Codex, Claude (incl. Claude Code), Cursor, Devin, Kilo, OpenCode, Cline, and related agent CLIs (e.g. Mimo, Jules). Agents share durable context via Ogham and Chroma, babysit PRs (CI + review threads), and work in isolated worktrees so parallel edits stay safe. Humans merge — agents prepare, never auto-merge.

Tooling: worktree-hive for issue → PR orchestration with isolated subagents.


🔬 Research Focus

  • SAAQ — Spiking Adaptive Activity Quantization: frontier MoE models on consumer GPUs
  • SNN compression — Low-bit, activity-aware quantization of large transformer MoE architectures
  • Symbolic regression — Compact equation discovery from high-dimensional GPU/neuromorphic telemetry
  • Explainable MoE — Dissect and visualize expert routing in open-weight models (no weight redistribution)

🛠️ Tech Stack

Actively learning and building with:

Languages:   Julia · Rust · Python · CUDA/C++ · HCL
Frameworks:  PyTorch · CUDA.jl · SymbolicRegression.jl · ort (ONNX)
Infra:       Terraform · Multi-cloud · NVIDIA Blackwell (RTX 5080)
Agents:      Grok Build · Codex · Claude · Cursor · Devin · Kilo · OpenCode · Cline
Memory:      Ogham · Chroma

📍 Currently

  • 🔥 SAAQ reference path via corinth-canal + Grok-scale quantization / dissect viz
  • 🧬 Growing the Limen Neural stack: encode → dynamics → NIR → FPGA
  • 🤖 Multi-agent PR and worktree workflows across the agent fleet above
  • ⚡ Local Blackwell / RTX 5080 kernels and training loops
  • 🎓 B.S. in AI Engineering @ WGU — deepening Julia, Rust, Python, CUDA, and HCL in real repos

README updated with Grok Build: Grok 4.5

Popular repositories Loading

  1. Ship-of-Theseus-HPC Ship-of-Theseus-HPC Public

    Localized HPC node for Bio-MEMS simulation, RTL design (SystemVerilog/Rust), and hardware diagnostics. Documentation for the 'Ship of Theseus' workstation.

  2. gaming-telemetry gaming-telemetry Public

    To record high demand output of DLSS 4.0 and path tracing as Neuromorphic data

    Rust

  3. corinth-canal corinth-canal Public

    Turning MOE architecture into SNN quantization

    Rust

  4. grok-ozempic grok-ozempic Public

    Turning all 318 GB of Grok 1 into SNN quantization, hence the name haha

    Rust

  5. Surrogate_Viz.jl Surrogate_Viz.jl Public

    Using SymbolicRegression.jl to create a new mathematical equation algorithm. Data that will be used is telemetry.csv from gaming-telemetry local repo.

    Julia

  6. skills-introduction-to-github skills-introduction-to-github Public

    My clone repository