mradermacher 发布 SEAD-SAGE-4B 的 GGUF 静态量化版本
mradermacher 为 SEAD-SAGE-4B 发布了 GGUF 静态量化版本,暂未提供 weighted/imatrix 量化。已提供从 Q2_K(1.9 GB)到 f16(8.9 GB)的多档量化,其中 Q4_K_S 与 Q4_K_M 标注为快速、推荐,Q6_K 为质量很好,Q8_0 为快速且最佳质量。
mradermacher 为 SEAD-SAGE-4B 发布了 GGUF 静态量化版本,暂未提供 weighted/imatrix 量化。已提供从 Q2_K(1.9 GB)到 f16(8.9 GB)的多档量化,其中 Q4_K_S 与 Q4_K_M 标注为快速、推荐,Q6_K 为质量很好,Q8_0 为快速且最佳质量。
Hugging Face 上线 yanggangu/CLIP-ViT-base-patch32-SMAT-DTD,这是 SMAT(Simple and Efficient Merge-Aware Training)论文表 2 中 seed 42 的专家模型,训练时即考虑模型合并。
Hugging Face 上线 yanggangu/CLIP-ViT-base-patch32-SMAT-SVHN,这是 SMAT(Simple and Efficient Merge-Aware Training)论文表 2 中针对 SVHN 的专家模型,seed 42。
Hugging Face 上线 CLIP-ViT-base-patch32-SMAT-RESISC45,这是 SMAT(Simple and Efficient Merge-Aware Training)论文表 2 中的专家模型,训练种子为 42。
Hugging Face 上线数据集 jetsonmom/omx_yellow_bear_box_2cam,由 LeRobot 生成,机器人类型为 omx_follower,含 20 条 episode、7043 帧、40 段视频,30 fps,动作与状态均为 6 自由度。
Hugging Face 上线 so-arm101-stack-green 数据集,由 LeRobot 生成,采用 codebase_version v3.0,机器人类型为 so_follower。数据集含 40 条 episode、32712 帧、1 个任务,30 fps,记录 6 自由度关节位置与前置、腕部两路 480×640 视频,全部划为训练集。
Hugging Face 上发布 strands-isaaclab-ant-131k-policy,这是用 rsl_rl PPO 在 Isaac Lab 的 Isaac-Ant 任务上训练的机器人策略,训练规模达 131,072 个并行环境。
Hugging Face 上线 nakanakagawa/sponge_task_9_29 数据集,由 LeRobot 生成,采用 v3.0 代码库版本,机器人类型为 so_follower。数据集含 4 条 episode、1596 帧、1 个任务,20 FPS,动作与状态均为 6 维关节位置,配 cam_0、cam_1 两路 480×640 视频。
Hugging Face Hub 上线 pigProfessional/act_so101_stack_green,这是一个用 LeRobot 0.6.2 训练的 ACT 模仿学习策略,机器人类型为 so_follower,输入 6 维状态与前置、腕部两路 480×640 图像,输出 6 维动作。
Hugging Face Hub 上出现了一个用 LeRobot 训练的 ACT(Action Chunking with Transformers)模仿学习策略模型 lenawngr/ACT_SWITCH-2-bottom_half1_test,采用 apache-2.0 许可。
strands-robots 通过 isaaclab train_policy provider 在 Isaac Lab 中训练宇树 Go2 四足机器人 PPO 策略,4096 个并行环境、1500 次迭代、147M 环境步,在单张 NVIDIA L40S 上耗时 57 分 39 秒,速度跟踪成功率 0.98。
推荐理由:原文给出从 Isaac Lab 训练 PPO 策略到录制 LeRobot v3 数据集的完整命令与验证流程,可迁移到其他机器人任务。
cagataydev 在 Hugging Face 发布 strands-isaaclab-shadow-reorient-policy,这是用 strands-robots 的 isaaclab train_policy provider(PR #4227)训练的 rsl_rl PPO 策略,任务为 Isaac-Reorient-Cube-Shadow 手中方块重定向。
推荐理由:读者可据此了解用 strands 工具链在 Isaac Lab 中训练并导出 Shadow Hand 转方块策略的完整流程与实测指标。
Hugging Face Hub 上线 klinzw/pi05_autolife_fridge,这是基于 Physical Intelligence π₀.₅、用 LeRobot 0.6.0 微调并推送的视觉语言动作策略,机器人本体为 autolife_s1,执行“open fridge”和“close fridge”任务。
一个基于 Vulkan 的 SLAM 系统正在开发中,可在 Android 和 Linux 上运行。项目目前更像概念验证,证明能在相机帧上以零拷贝方式运行任意计算着色器,而非实际的 SLAM/VIO,作者计划未来实现后者。
开发者 haihoang1219931 开源了 SenseRobot AI 国际象棋训练器的 DIY 克隆方案,用 Arduino Mega 2560 + RAMP 1.5 和 NEMA 步进电机复现其物理移子机构、AI 视觉追踪与机械臂精度,无需 ROS。
Linorobot2 Cockpit 发布 v2026.09,这是一个容器化、基于浏览器的 ROS 2 移动机器人监管器,覆盖 micro-ROS 固件生成、硬件验证、物理仿真调参与自主导航。
ArticuRiddle 是一个包含 100 个铰接物体的数据集,用于测试视觉外观与真实关节运动相矛盾时的操作能力。每个物体由 PartManip 训练资产经两类编辑生成:50 个移动或旋转把手使其暗示错误运动,50 个直接交换关节类型(如抽屉面板改为摆动、门改为滑动),外观保持不变。
Hugging Face 上线数据集 Sterzhang/video2skill-bench,页面仅保留“以开源和开放科学推进 AI 民主化”的使命说明,未披露数据规模、任务类型或评测指标等细节。
QERRA-v2 Classical 发布 v2.0.2,为 Layer 3(QERRA-THRIVE)全部 12 个行为向量加入 40 字符回看否定守卫,修复安全候选因表述规避风险而被扣分导致的排序反转。
Hugging Face 上的 Mulligan 项目发布 sim-square-narrow-r03-mining-no-cf-divl,这是一个基于父代冻结扩散 actor 与分布式 DIVL critic 的状态型智能体,训练步数 150001,含 5 个随机种子检查点。
Hugging Face 在 Mulligan 项目下发布基于状态的 IDQL 智能体检查点 sim-square-narrow-r03,采用扩散 actor 加标量 IQL critic,训练步数 150001,覆盖 5 个随机种子。
Hugging Face 上的 Mulligan 项目发布 sim-square-narrow-r03-mulligan-divl 检查点,基于父代冻结扩散 actor 与分布式 DIVL critic,在 sim-square-narrow 任务训练至 150001 步,含 5 个随机种子。
Hugging Face 的 Mulligan 项目发布 sim-square-narrow-r03-mining-no-cf-idql,这是一个基于状态的 IDQL 智能体,采用扩散 actor 与标量 IQL critic,包含 policy.pt 和 stats.json 归一化器。
Hugging Face 上的 Mulligan 项目发布 sim-square-narrow-r03-baseline-repair-r3roll-idql,是基于状态的 IDQL 智能体,含扩散 actor 与标量 IQL critic,提供 policy.pt 和 stats.json。
Hugging Face 上的 Mulligan 项目发布 sim-square-narrow-r03-baseline-repair-r3roll-divl 检查点,基于父代冻结扩散 actor 与分布式 DIVL critic,训练至 150001 步,含 5 个随机种子。
Mulligan 在 Hugging Face 发布基于状态的 IDQL 智能体 sim-square-narrow-r03-auto-plain-il-n1-idql,用于 sim-square-narrow 任务,包含 5 个随机种子、训练步数 150001 的 PyTorch 检查点。
Hugging Face 上的 Mulligan 组织发布 sim-square-narrow-r03-auto-iql-success-bc-n32-idql,这是一个基于状态的 IDQL 智能体,采用扩散 actor 与标量 IQL critic,任务为 sim-square-narrow,训练步数 150001,含 5 个随机种子检查点。
Hugging Face 上的 Mulligan 组织发布 sim-square-narrow-r03-auto-iql-success-bc-n32-divl 检查点,这是一个基于状态的智能体,使用父级冻结扩散 actor 与分布式 DIVL critic,任务为 sim-square-narrow,训练步数 150001,含 5 个种子。
Hugging Face 上 Mulligan 组织发布 sim-square-narrow-r03-auto-iql-n32-idql,这是一个基于状态的 IDQL 智能体,采用扩散 actor 与标量 IQL critic,任务为 sim-square-narrow,属 R3 轮次。
Hugging Face 上线 mulligan/sim-square-narrow-r03-auto-iql-n32-divl 模型仓库。该条目名称显示其与 sim-square-narrow 任务、auto-iql 方法及 n32 配置相关,具体性能与训练细节尚未在页面中披露。
Hugging Face 上线了名为 mulligan/sim-square-narrow-r03-auto-filtered-bc-n1-idql 的模型仓库。该页面未披露模型架构、参数规模、训练数据或评测结果等具体信息,仅以推进开源开放科学为宗旨说明。
Hugging Face 上的 Mulligan 项目发布 sim-square-narrow-r02-mining-no-cf-idql,这是一个基于状态的 IDQL 智能体,由扩散 actor 与标量 IQL critic 组成,提供 policy.pt 和 stats.json。
Hugging Face 上的 Mulligan 项目发布 sim-square-narrow-r02-mining-no-cf-divl,这是一个基于状态、使用父级冻结扩散 actor 与分布式 DIVL critic 的智能体,任务为 sim-square-narrow,训练步数 150001,含 1-5 共 5 个种子检查点。
Hugging Face 上的 Mulligan 项目发布 sim-square-narrow-r02 的 IDQL 智能体检查点,采用扩散 actor 加标量 IQL critic,训练步数 150001,覆盖 5 个随机种子。
Hugging Face 上的 Mulligan 项目发布 sim-square-narrow-r03-baseline-idql,为基于状态的 IDQL 智能体,采用扩散 actor 与标量 IQL critic,含 policy.pt 和 stats.json。
Hugging Face 上的 Mulligan 项目发布 sim-square-narrow-r02-baseline-divl 基线智能体,用于 sim-square-narrow 任务,采用父级冻结扩散 actor 与分布式 DIVL critic,训练步数 150001,含 5 个随机种子检查点。
Hugging Face 的 Mulligan 项目发布 sim-square-narrow-r03-baseline-divl,这是一个基于状态的智能体,使用父级冻结扩散 actor 与分布式 DIVL critic,任务为 sim-square-narrow,训练步数 150001。
Hugging Face 上的 Mulligan 项目发布 sim-square-narrow-r02-baseline-idql 基线检查点,为基于状态的 IDQL 智能体,采用扩散 actor 与标量 IQL critic,含 policy.pt 和 stats.json。
Hugging Face 的 Mulligan 项目发布 sim-square-narrow-r03-mulligan-repair-no-r3roll-divl 检查点,基于父代冻结扩散 actor 与分布式 DIVL 评论家,训练步数 150001,含 1-5 共 5 个种子。
Hugging Face 上的 Mulligan 项目发布 sim-square-narrow-r03-mulligan-repair-no-r3roll-idql 检查点,为基于状态的 IDQL 智能体,采用扩散 actor 加标量 IQL critic,含 policy.pt 与 stats.json 归一化文件。