ZJUNLP avatar

ZJUNLP

Official

@zjunlp · China

0Followers
|
108Public Repos
|
222Published Skills

Knowledge Engine Lab: A NLP & KG Group of Zhejiang University

Skills Distribution
DomainAI Models & ...Simulated Environm.. (40%)E-commerce Decisio.. (30%)Structured Data An.. (30%)

Agent Skills by ZJUNLP

Showing 222 vetted skills indexed across 3 GitHub repositories.

zjunlpzjunlp
75

auto-verify

Stress-test research claims by swapping method, dataset, and model variants.

Official
Advanced
zjunlpzjunlp
75

impact-check

Assess the importance and reach of a research idea before committing effort.

Official
Intermediate
zjunlpzjunlp
75

notify

Drafts and dispatches research-progress briefings through user-configured notification services.

Official
Intermediate
zjunlpzjunlp
75

mechanism-behavior-discovery

Surfaces novel falsifiable behavioral phenomena in LLMs for downstream mechanistic investigation.

Official
Intermediate
zjunlpzjunlp
75

training-check

Monitors WandB training metrics periodically to detect NaN, divergence, and stalled runs.

Official
Intermediate
zjunlpzjunlp
75

mechanism-skills

Route mechanistic interpretability questions to eleven method families for localizing internal model components.

Official
Advanced
zjunlpzjunlp
75

result-to-claim

Evaluates experiment results against intended claims using an external LLM reviewer and routes next actions.

Official
Advanced
zjunlpzjunlp
75

hypothesis-batch

Generates and refines batches of research hypotheses through a multi-phase automated pipeline.

Official
Advanced
zjunlpzjunlp
75

auto

Orchestrates autonomous research pipelines from claim generation through experiments to verification and iteration.

Official
Advanced
zjunlpzjunlp
75

research-refine-pipeline

Chains method refinement and experiment planning into one end-to-end research proposal workflow.

Official
Advanced
zjunlpzjunlp
75

experiment-audit

Audits per-claim experimental methodology integrity using cross-model LLM review.

Official
Advanced
zjunlpzjunlp
75

experiment-plan

Converts a refined research proposal into a claim-driven experiment roadmap with run order and budgets.

Official
Advanced
zjunlpzjunlp
75

idea-creator

Generate, validate, and rank research ideas with pilot experiments for a given direction.

Official
Advanced
zjunlpzjunlp
75

experiment-queue

Orchestrates batched ML experiments on SSH GPU servers with OOM retry and wave scheduling.

Official
Advanced
zjunlpzjunlp
75

mhistory

Generates a chronological research development-history article from database retrieval and web search.

Official
Advanced
zjunlpzjunlp
75

mechanic-db-search

Retrieves academic papers from interpretability and cross-disciplinary databases via a cloud search service.

Official
Advanced
zjunlpzjunlp
75

monitor-experiment

Monitor remote GPU experiments, collect results, and finalize cost manifests over SSH.

Official
Advanced
zjunlpzjunlp
75

research-refine

Refines vague research directions into concrete method proposals via iterative external LLM review.

Official
Advanced
zjunlpzjunlp
75

data-rule

Defines dataset provenance, split, label, and sample-size constraints for experiments.

Official
Basic
zjunlpzjunlp
75

ablation-planner

Designs and runs ablation studies for ML experiments using an external LLM reviewer.

Official
Advanced
zjunlpzjunlp
75

auto-claim

Orchestrates the claim-stage pipeline producing research proposals and experiment plans.

Official
Advanced
zjunlpzjunlp
75

mechanism-audit

Audit mechanistic interpretability experiment rigor per claim using cross-model LLM review.

Official
Advanced
zjunlpzjunlp
75

analyze-results

Analyze ML experiment results and generate comparison tables with statistical insights.

Official
Intermediate
zjunlpzjunlp
75

paper-figure

Generate publication-quality matplotlib figures and LaTeX tables from experiment data files.

Official
Advanced

Frequently Asked Questions About ZJUNLP

FAQPage Schema
What specific tasks can I perform using ZJUNLP's simulated environment skills?▼

You can execute complex object manipulation tasks, including locating, heating, cooling, cleaning, and storing items within ALFWorld and ScienceWorld environments, as well as conducting scientific experiments like circuit building and substance mixing.

Which personas benefit most from these technical capabilities?▼

Researchers and developers focused on embodied cognition, decision-making benchmarks, and multi-turn data analysis will find these capabilities most relevant for testing and refining complex reasoning systems.

What are the primary data dependencies for the financial and clinical analysis skills?▼

These skills require structured datasets, specifically SQLite databases containing MIMIC-IV patient records or SEC 10-K financial filings, to perform metric extraction and report generation.