WO2020118730A1 - 机器人柔顺性控制方法、装置、设备及存储介质 - Google Patents
机器人柔顺性控制方法、装置、设备及存储介质 Download PDFInfo
- Publication number
- WO2020118730A1 WO2020118730A1 PCT/CN2018/121338 CN2018121338W WO2020118730A1 WO 2020118730 A1 WO2020118730 A1 WO 2020118730A1 CN 2018121338 W CN2018121338 W CN 2018121338W WO 2020118730 A1 WO2020118730 A1 WO 2020118730A1
- Authority
- WO
- WIPO (PCT)
- Prior art keywords
- data
- teaching
- motion
- variable impedance
- movement
- Prior art date
- Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
- Ceased
Links
Images
Classifications
-
- G—PHYSICS
- G05—CONTROLLING; REGULATING
- G05B—CONTROL OR REGULATING SYSTEMS IN GENERAL; FUNCTIONAL ELEMENTS OF SUCH SYSTEMS; MONITORING OR TESTING ARRANGEMENTS FOR SUCH SYSTEMS OR ELEMENTS
- G05B19/00—Program-control systems
- G05B19/02—Program-control systems electric
- G05B19/42—Recording and playback systems, i.e. in which the program is recorded from a cycle of operations, e.g. the cycle of operations being manually controlled, after which this record is played back on the same machine
- G05B19/423—Teaching successive positions by walk-through, i.e. the tool head or end effector being grasped and guided directly, with or without servo-assistance, to follow a path
-
- G—PHYSICS
- G05—CONTROLLING; REGULATING
- G05B—CONTROL OR REGULATING SYSTEMS IN GENERAL; FUNCTIONAL ELEMENTS OF SUCH SYSTEMS; MONITORING OR TESTING ARRANGEMENTS FOR SUCH SYSTEMS OR ELEMENTS
- G05B2219/00—Program-control systems
- G05B2219/30—Nc systems
- G05B2219/39—Robotics, robotics to robotics hand
- G05B2219/39319—Force control, force as reference, active compliance
Definitions
- the invention belongs to the field of computer technology, and particularly relates to a robot compliance control method, device, equipment and storage medium.
- the trajectory of the robotic arm is generally pre-defined by the user, or a certain task environment is preset, and then the robot or the robotic arm can be repeatedly executed according to the plan.
- the robotic arm operating in this mode cannot face environmental changes or sudden disturbances.
- this mode also requires more arduous manual programming.
- the use threshold is high (for example: to be able to program robots). More importantly, this robot control mode does not imply human operation habits, nor is it as flexible as human hands.
- the robotic arm or robot should have learning capabilities and be more flexible and compliant.
- the robot "Imitation Learning” (Imitation Learning) or “Teaching Learning” (Programming by Demonstration) is an important method to solve this problem.
- the compliant behavior of a robot includes two aspects of action and force, so the learning of compliant behavior also includes two aspects of action learning and force learning.
- the existing robot compliance control method independently models and learns the motion trajectory and force, and the learning effect is not good, which leads to inaccurate control results; based on Gaussian mixture model, Gaussian process and other offline regression methods to For imitation learning, the training time required is relatively long, and the training efficiency is low; the stability of the control cannot be guaranteed, and there may be situations where the robot interaction force is too large and hurts people.
- the object of the present invention is to provide a robot compliance control method, device, equipment and storage medium, aiming to solve the problems of inaccurate control results and poor control effects caused by poor compliance of the existing robot compliance control methods.
- the present invention provides a robot compliance control method, which includes the following steps:
- the teaching data includes at least the movement data and interaction force data of the teaching movement
- variable impedance Calculating the motion equation of the teaching motion based on the motion data in the teaching data, and simultaneously calculating the variable impedance parameter of the teaching motion based on the interaction force data in the teaching data, wherein the variable impedance
- the parameters include at least variable stiffness parameters and variable damping parameters
- the operation is controlled according to the equation of motion and the variable impedance parameter.
- the present invention provides a robot compliance control device, the device including:
- a data acquisition unit for acquiring teaching data of the teaching movement, wherein the teaching data includes at least the movement data and the interaction force data of the teaching movement;
- a parameter calculation unit configured to calculate the motion equation of the teaching motion based on the motion data in the teaching data, and at the same time calculate the variable impedance parameter of the teaching motion based on the interaction force data in the teaching data,
- the variable impedance parameter includes at least a variable stiffness parameter and a variable damping parameter;
- the operation control unit is used for controlling operation according to the motion equation and the variable impedance parameter.
- the present invention also provides a computing device, including a memory, a processor, and a computer program stored in the memory and executable on the processor, which is implemented when the processor executes the computer program The steps of the robot compliance control method as described.
- the present invention also provides a computer-readable storage medium that stores a computer program, and when the computer program is executed by a processor, implements the steps of the robot compliance control method.
- the invention obtains the teaching data of the teaching movement, calculates the motion equation of the teaching movement based on the movement data in the teaching data, and simultaneously calculates the teaching according to the interaction force data in the teaching data
- the variable impedance parameter of the movement controls the operation according to the motion equation and the variable impedance parameter, thereby reducing manual programming during the robot compliance control process, lowering the threshold for robot use, and improving the flexibility and accuracy of robot control , And further improve the robot's generalization ability, intelligence and control effect.
- Embodiment 1 is an implementation flowchart of a robot compliance control method provided in Embodiment 1 of the present invention
- FIG. 2 is a schematic structural diagram of a robot compliance control device according to Embodiment 2 of the present invention.
- FIG. 3 is a schematic structural diagram of a robot compliance control device according to Embodiment 3 of the present invention.
- FIG. 4 is a schematic structural diagram of a computing device according to Embodiment 4 of the present invention.
- FIG. 5 is a schematic structural diagram of a robot compliance control device according to Embodiment 2 of the present invention.
- FIG. 6 is a schematic structural diagram of a computing device according to Embodiment 3 of the present invention.
- FIG. 1 shows the implementation flow of the robot compliance control method provided in Embodiment 1 of the present invention. For convenience of description, only the parts related to the embodiment of the present invention are shown. The details are as follows:
- step S101 the teaching data of the teaching movement is acquired.
- the embodiments of the present invention are suitable for automatic control of robots.
- Robots include a series of robot products that are not limited to robotic arms, humanoid robots, etc. with joints, links, and other structures, and can achieve telescopic and grasping actions.
- the teaching data may include at least motion data and interaction force data of the teaching movement. Therefore, learning of the teaching movement may include action learning and force learning (ie, variable stiffness parameter and variable damping parameter learning).
- the motion data may include position data and speed data of a preset point of the robot (eg, end effector), or the motion data may include angle and angle of a preset angle of the robot (eg, joint angle) Acceleration.
- the motion data may also include one or more other parameters that can be used to completely describe the teaching motion, which is not limited by the present invention.
- FIG. 2 shows a diagram of teaching a robot.
- the teacher grasps the end effector of the robot with one hand and moves in a plane or space. Make a trajectory, and the other hand exerts a teaching force at the end.
- the robot collects teaching data through its own motion capture system and a six-dimensional force sensor mounted on the wrist.
- the motion data includes position data and velocity data of a preset sampling point of the robot (eg, end effector or end, etc.)
- the position data and interaction force data of the end are sampled at time intervals to obtain a series of samples
- the teacher controls the robot through the remote control or the teach pendant to perform the teaching operation, or teaches by hand.
- the robot records the teaching data according to the teaching operation.
- the instructor personally completes the teaching movement task.
- the teaching data is collected by the robot's motion catcher, data glove, and force sensor according to the teaching movement.
- the motion data includes position data and speed data of a preset point of the robot (for example, an end effector)
- the position data related to the teaching motion may be obtained first, Interaction force data and time data, and then calculate the speed data related to the teaching movement based on the position data and the time data, thereby obtaining the movement data of the teaching movement.
- the motion data includes the angle and angular acceleration of the preset angle of the robot (for example, the joint angle)
- the angle data and interaction force related to the teaching motion may be obtained first Data and time data, and then calculate the angular acceleration data related to the teaching movement based on the angle data and the time data, thereby obtaining the movement data of the teaching movement.
- step S102 the motion equation of the teaching motion is calculated based on the motion data in the teaching data, and at the same time, the variable impedance parameter of the teaching motion is calculated based on the interaction force data in the teaching data.
- variable impedance parameter may include at least a variable stiffness parameter and a variable damping parameter.
- trajectory and force are learned, so as to improve the learning effect and thus the accuracy of the control results.
- the preset neural network model can be trained using the motion data to obtain the motion equation of the teaching motion, and the neural network can be trained according to the motion equation
- the model is updated online, thereby improving the calculation efficiency of the motion equation, facilitating the subsequent use of the motion equation, and adapting to the needs of real-time online learning, thereby improving the learning effect.
- the motion data when using the motion data to train the preset neural network model, can be incrementally learned one by one or block by block to obtain the motion equation of the teaching motion, thereby improving the accuracy of the motion equation Sex, thereby improving learning effectiveness.
- the neural network model can be a support vector machine (Support Vector Machine, SVM), online sequence over-limit learning machine and other models that can be incrementally online learning, or other incremental online learning models, such as incremental support vector machine (ISVM), etc., the present invention does not limit this.
- SVM Support Vector Machine
- the online sequence over-limit learning machine has the characteristics of fast learning speed, strong generalization ability, and simple implementation. Therefore, preferably, the neural network model is an online sequence over-limit learning machine, that is, use The motion data trains the online sequence overrun learning machine, thereby improving the training efficiency.
- the input and output are the position and speed (or the angle and angular acceleration of the joint angle) of the sampling point (for example, the robot end effector), so the online
- FIG. 3 shows an exemplary structure of an online sequence overrun learning machine.
- the activation function of the hidden layer of the online sequence overrun learning machine is g
- the online sequence overrun learning we want to learn
- the machine ie, the model to be learned
- the number of hidden layer neurons is For the offset of the hidden layer, Is the weight of the hidden layer, the dimension is Is the weight of the output layer, the dimension is
- W and b are randomly generated and fixed.
- the training process only needs to determine the weight of the output layer. Optimization process to achieve.
- the activation function g generally selects the sigmoid function (sigmoid function) or the hyperbolic tangent function (tanh function), the modified sigmoid function can also be used, for example, However, as long as it is satisfied And the monotonically increasing continuous and continuously differentiable functions all meet the requirements of the activation function, and are not limited here.
- the training goal of the online sequence overrun learning machine is to find a set of optimal output layer weights
- H + is the Moore-Penrose generalized inverse matrix of the matrix H.
- the output layer weights can be obtained without iteration.
- the constraints are added, the problem of solving the output layer weights becomes a constrained optimization problem.
- the training process of the online sequence overrun learning machine includes an initial ELM batch learning process and a continuous sequential learning process, as follows:
- N 1 is the newly arrived data, by The calculated initial output weight is among them, Whenever a new training sample is obtained When Recursively calculate output weights. among them,
- variable stiffness parameter and the variable damping parameter may be calculated based on the interaction force data.
- variable stiffness parameter when calculating the variable stiffness parameter, let Represents the collected interaction force (F) and corresponding time (q) information, where is the number of disturbance data samples obtained.
- the variable stiffness parameter at time q is calculated from the force information in the time window [q-(w-1), q].
- the length of the sliding time window is w, and the upper and lower bounds of the data points in the window are represented by L q and U q ,
- the stiffness matrix K q is among them, And eigenvalues Proportional to the expression
- the interactive force data will be continuously collected, and the new data will be sorted according to the time information and the values in the window will be taken to solve the stiffness. For example, when the data at time q+1 enters, the online update of the covariance is among them,
- variable damping parameter B since the damping ratio is constant, the square root of the damping and the stiffness is linear, so it can be based on the formula To calculate the variable damping parameter B. Among them, ⁇ is a constant greater than 0.
- the preset stability constraints and the interaction force data can be used to predict
- the variable impedance model is trained to obtain the variable impedance parameters of the teaching movement, and the variable impedance model is updated according to the variable impedance parameters, so as to ensure the stability of the variable impedance control and avoid excessive robot interaction force that may cause injury. happening.
- step S103 the operation is controlled according to the equation of motion and the variable impedance parameter.
- the operation can be controlled according to the motion equation and the variable impedance parameter, thereby controlling the robot to reproduce the movement trajectory and interactive force of the teaching movement.
- the trained neural network model for example, online sequence overrun learning machine
- variable impedance models to control the trajectory and interaction force of the robot to reproduce the teaching movement.
- FIG. 4 shows an exemplary diagram of teaching learning and reproduction of robot compliance control.
- the instructor grasps the robot with one hand for teaching, and the robot collects Track information And force information F q , then according to the trajectory information Perform motion learning to obtain f( ⁇ ), and learn variable stiffness parameters and variable damping parameters according to the force information F q to get ⁇ B q ,K q ⁇ , and finally generate motion according to f( ⁇ ) and according to ⁇ B q ,K q ⁇ Variable impedance control to control the trajectory and interaction force of the robot to reproduce the teaching movement.
- the motion equation of the teaching movement is calculated according to the movement data in the teaching data, and at the same time, the variation of the teaching movement is calculated according to the interaction force data in the teaching data Impedance parameters, according to the motion equation and variable impedance parameter control operation, thereby reducing the manual programming in the robot compliance control process, lowering the threshold for robot use, improving the flexibility and accuracy of robot control, and thus improving the robot's universal Ability, degree of intelligence and control effect.
- FIG. 5 shows the structure of the robot compliance control device provided in Embodiment 2 of the present invention. For ease of explanation, only parts related to the embodiment of the present invention are shown, including: a data acquisition unit 51, a parameter calculation unit 52 and Operation control unit 53.
- the data acquiring unit 51 is configured to acquire teaching data of the teaching movement, wherein the teaching data includes at least the movement data and the interaction force data of the teaching movement.
- the learning of the teaching movement may include action learning and force learning (ie, variable stiffness parameter and variable damping parameter learning).
- the motion data may include position data and speed data of a preset point of the robot (eg, end effector), or the motion data may include angle and angle of a preset angle of the robot (eg, joint angle) Acceleration.
- the motion data may also include one or more other parameters that can be used to completely describe the teaching motion, which is not limited by the present invention.
- the data acquisition unit 51 may include:
- the first acquiring unit is used to acquire position data, interaction force data and time data related to the teaching movement
- the first calculation unit is used to calculate the motion data according to the position data and the time data.
- the speed data related to the teaching movement is calculated based on the position data and the time data, thereby obtaining the movement data of the teaching movement.
- the data acquisition unit 51 may further include:
- the second acquisition unit is used to acquire angle data, interaction force data and time data related to the teaching movement.
- the second calculation unit is used to calculate the motion data according to the angle data and the time data.
- the angular acceleration data related to the teaching movement may be calculated according to the angle data and the time data, thereby obtaining the movement data of the teaching movement.
- the parameter calculation unit 52 is configured to calculate the motion equation of the teaching motion based on the motion data in the teaching data, and at the same time calculate the variable impedance parameter of the teaching motion based on the interaction force data in the teaching data, where the variable impedance parameter includes at least Variable stiffness parameters and variable damping parameters.
- the trajectory and the force are simultaneously learned, thereby improving the learning effect, and thereby improving the accuracy of the control result.
- the parameter calculation unit 52 may include:
- the first training unit is used to train the preset neural network model using motion data to obtain the motion equation of the teaching movement, and update the neural network model online according to the motion equation, thereby improving the calculation efficiency of the motion equation and facilitating movement
- the model training unit may include:
- the incremental learning unit is used to incrementally learn the motion data in a one-by-one or block-by-block manner to obtain the motion equation of the teaching motion, thereby improving the accuracy of the motion equation and thereby improving the learning effect.
- the neural network model is an online sequence overrun learning machine.
- the parameter calculation unit 52 may include:
- the second training unit is used to train the preset variable impedance model according to the preset stability constraints and interaction force data to obtain the variable impedance parameter of the teaching movement, and update the variable impedance model according to the variable impedance parameter, Therefore, the stability of the variable impedance control is ensured, and the situation that the interaction force of the robot is too large to cause injury is avoided.
- the operation control unit 53 is used to control the operation according to the equation of motion and the variable impedance parameter.
- the teaching data of the teaching movement is acquired by the data acquiring unit 51, the motion equation of the teaching movement is calculated according to the movement data in the teaching data by the parameter calculating unit 52, and at the same time according to the teaching data
- the interactive force data calculates the variable impedance parameters of the teaching movement, and the operation is controlled by the operation control unit 53 according to the motion equation and the variable impedance parameters, thereby reducing the manual programming in the robot compliance control process, lowering the robot's use threshold, and improving the robot The flexibility and accuracy of the control, thereby improving the robot's generalization ability, intelligence and control effect.
- each unit of the robot compliance control device may be implemented by a corresponding hardware or software unit, and each unit may be an independent software and hardware unit, or may be integrated into one software and hardware unit, which is not limited here. this invention.
- FIG. 6 shows the structure of the computing device provided in Embodiment 4 of the present invention. For ease of description, only parts related to the embodiment of the present invention are shown.
- the computing device 6 of the embodiment of the present invention includes a processor 60, a memory 61, and a computer program 62 stored in the memory 61 and executable on the processor 60.
- the processor 60 executes the computer program 62
- the steps in the above embodiments of the robot compliance control method are implemented, for example, steps S101 to S103 shown in FIG. 1.
- the processor 60 executes the computer program 62
- the functions of the units in the above device embodiments are realized, for example, the functions of the units 51 to 53 shown in FIG.
- the teaching data of the teaching motion is acquired, and the teaching data is calculated according to the motion data in the teaching data Teaching the motion equation of motion, and at the same time calculating the variable impedance parameters of the teaching motion based on the interactive force data in the teaching data, and controlling the operation according to the motion equations and variable impedance parameters, thereby reducing the manual programming and reducing the robot's compliance control process.
- the use threshold of the robot is improved, and the flexibility and accuracy of the robot control are improved, thereby improving the robot's generalization ability, intelligence, and control effect.
- a computer-readable storage medium stores a computer program, and when the computer program is executed by a processor, the steps in the embodiments of the foregoing robot compliance control methods are implemented. For example, steps S101 to S103 shown in FIG. 1.
- the functions of the units in the foregoing device embodiments are realized, for example, the functions of the units 51 to 53 shown in FIG. 5.
- the teaching data of the teaching motion is acquired, the motion equation of the teaching motion is calculated according to the motion data in the teaching data, and the variable impedance of the teaching motion is calculated according to the interaction force data in the teaching data
- the parameters control the operation according to the equations of motion and variable impedance parameters, thereby reducing manual programming during robot compliance control, lowering the threshold for robot use, improving the flexibility and accuracy of robot control, and thus improving the generalization of the robot Ability, degree of intelligence and control effect.
- the robot compliance control method implemented when the computer program is executed by the processor may further refer to the description of the steps in the foregoing method embodiments, and details are not described herein again.
- the computer-readable storage medium in the embodiments of the present invention may include any entity or device capable of carrying computer program code, and a recording medium, such as ROM/RAM, magnetic disk, optical disk, flash memory, and other memories.
Landscapes
- Engineering & Computer Science (AREA)
- Robotics (AREA)
- Physics & Mathematics (AREA)
- General Physics & Mathematics (AREA)
- Automation & Control Theory (AREA)
- Manipulator (AREA)
Abstract
一种机器人柔顺性控制方法,该方法包括:获取示教运动的示教数据,根据示教数据中的运动数据计算示教运动的运动方程,并且同时根据示教数据中的交互力数据计算示教运动的变阻抗参数,根据运动方程和变阻抗参数控制操作,从而减少了机器人柔顺性控制过程中的手动编程,降低了机器人的使用门槛,提高了机器人控制的柔顺性和精确性,进而提高了机器人的泛化能力、智能化程度和控制效果。还涉及机器人柔顺性控制装置、设备及存储介质。
Description
本发明属于计算机技术领域,尤其涉及一种机器人柔顺性控制方法、装置、设备及存储介质。
在现阶段机器人的应用中,尤其是工业应用中,机械臂的运动轨迹一般是通过用户预先定义的,或者预先设定某种任务环境,然后让机器人或机械臂按照计划重复执行即可。这种模式运行的机械臂无法面对环境的变化,或者突如其来的扰动。对于复杂场景下或较困难任务的实现,这种模式也需要较为繁重的人工编程。对普通工人来讲,使用门槛要求高(例如:要会机器人编程)。更重要的是,这种机器人控制模式没有隐含人的操作习惯更没有像人手那样的具有柔顺性。为了有效降低机器人的使用门槛、更好地实现人机协同交互,机械臂或机器人应该具有学习能力,并更加灵活和柔顺的特性。机器人“模仿学习”(Imitation Learning)或者“示教学习”(Programming by Demonstration)便是解决这一问题的重要方法。
通常机器人的柔顺性行为包含动作和力两个方面,因此柔顺性行为的学习也是包括动作学习和力学习两个方面的。
在机器人柔顺性控制领域,之前的研究工作主要集中在控制器的人为设计领域(例如:力位混合控制,阻抗控制,碰撞检测反馈控制器等)以及被动柔顺机构设计。上述的柔顺性控制器设计方法具有复杂的调参过程,且不具有泛化能力不能适应新的情况。机器人通过学习人类柔顺性行为而获得柔顺控制策略的研究能简化复杂的调参过程和减低机器人的使用门槛(工人只需提供正确的人类示教即可让机器人具有相应的柔顺性行为,而不需要使用者具有编程和 机器人控制的相关技术基础)。
机器人通过学习人类柔顺性行为而获得柔顺控制策略的研究属于前沿领域,示教学习控制中大都将运动轨迹学习和力的学习独立研究。例如,Seyed Mohammad Khansari-Zadeh提出一种学习运动轨迹的方法(发表于2011年IEEE Transactions on Robotics上的文章《Learning Stable Nonlinear Dynamical Systems With Gaussian Mixture Models》)。在该方法的最初提出时,动态系统是通过高斯混合模型(Gaussian Mixture Models)来建模的,并且基于李雅普诺夫稳定性的约束也被推导出来用于保证运动收敛到目标。在接下来几年的发展中也出现了其他一些轨迹动作的模仿学习方法,但使用动态系统建模、使用李雅普诺夫稳定性进行约束这两大特征,基本上是各种方法的共同特征。Calinon提出了一种根据示教位置扰动的协方差推导出不同交互力的学习方法,但是这种方法示教怪异且不利于与轨迹一起学习。
从现有的资料来看,将运动轨迹和力看做柔顺性行为两个组成部分,并将两者用于机器人柔顺性行为学习控制的相关成熟方案甚少。2017年发表于Autonomous Robots的文章《Learning potential functions from human demonstrations with encapsulated dynamic and compliant behaviors》提出了基于势函数和耗散场的联合变阻抗控制策略,该方法的需要人为通过先验知识设计多组基于任务的参数,这种方法具有强构造性,且只能离线训练无法,效率低。
已申请或已授权的专利中,也有一些与所述领域有关。在名称为“一种基于高斯过程的机器人模仿”的专利文件中,公开了一种基于高斯过程的机器人模仿学习方法。高斯过程也是一种回归算法,与高斯混合模型类似,该方案使用高斯过程对机器人运动进行建模学习。在名称为“一种基于轨迹模仿的机器人汉字书写学习方法”的专利文件中,公开了一种将基于轨迹匹配的模仿学习引入到机器人书写技能的学习中,将汉字的比划进行分割,并通过多个高斯混合模型对示教数据进行编码学习和重构的方法。在名称为“具有模仿学习机制的手把手示教机械臂系统及方法”的专利文件中,公开了一种带有模仿学习功 能的机械臂系统,并给出了基于前馈神经网络的模仿学习建模方法。在名称为“一种机器人力控示教模仿学习的装置及方法”的专利文件中,公开了一种在示教数据中引入了力反馈信息,并使用隐马尔科夫模型对示教数据进行建模编码的方法。
综上所述,现有的机器人柔顺性控制方法,对运动轨迹和力进行独立建模学习,学习效果不佳,进而导致控制结果不精确;基于高斯混合模型、高斯过程等离线的回归方法来进行模仿学习,需要的训练时间比较长,训练效率较低;控制的稳定性无法保证,可能出现机器人交互力过大而出现伤人的情况。
发明内容
本发明的目的在于提供一种机器人柔顺性控制方法、装置、设备及存储介质,旨在解决现有的机器人柔顺性控制方法的控制结果不精确、柔顺性较差导致的控制效果不佳问题。
一方面,本发明提供了一种机器人柔顺性控制方法,所述方法包括下述步骤:
获取示教运动的示教数据,其中,所述示教数据至少包括所述示教运动的运动数据和交互力数据;
根据所述示教数据中的运动数据计算所述示教运动的运动方程,并且同时根据所述示教数据中的交互力数据计算所述示教运动的变阻抗参数,其中,所述变阻抗参数至少包括变刚度参数和变阻尼参数;
根据所述运动方程和所述变阻抗参数控制操作。
另一方面,本发明提供了一种机器人柔顺性控制装置,所述装置包括:
数据获取单元,用于获取示教运动的示教数据,其中,所述示教数据至少包括所述示教运动的运动数据和交互力数据;
参数计算单元,用于根据所述示教数据中的运动数据计算所述示教运动的运动方程,并且同时根据所述示教数据中的交互力数据计算所述示教运动的变 阻抗参数,其中,所述变阻抗参数至少包括变刚度参数和变阻尼参数;以及
操作控制单元,用于根据所述运动方程和所述变阻抗参数控制操作。
另一方面,本发明还提供了一种计算设备,包括存储器、处理器以及存储在所述存储器中并可在所述处理器上运行的计算机程序,所述处理器执行所述计算机程序时实现如所述机器人柔顺性控制方法的步骤。
另一方面,本发明还提供了一种计算机可读存储介质,所述计算机可读存储介质存储有计算机程序,所述计算机程序被处理器执行时实现如所述机器人柔顺性控制方法的步骤。
本发明通过获取示教运动的示教数据,根据所述示教数据中的运动数据计算所述示教运动的运动方程,并且同时根据所述示教数据中的交互力数据计算所述示教运动的变阻抗参数,根据所述运动方程和所述变阻抗参数控制操作,从而减少了机器人柔顺性控制过程中的手动编程,降低了机器人的使用门槛,提高了机器人控制的柔顺性和精确性,进而提高了机器人的泛化能力、智能化程度和控制效果。
图1是本发明实施例一提供的机器人柔顺性控制方法的实现流程图;
图2是本发明实施例二提供的机器人柔顺性控制装置的结构示意图;
图3是本发明实施例三提供的机器人柔顺性控制装置的结构示意图;以及
图4是本发明实施例四提供的计算设备的结构示意图。
[根据细则91更正 01.01.2019]
图5是本发明实施例二提供的机器人柔顺性控制装置的结构示意图;以及
图5是本发明实施例二提供的机器人柔顺性控制装置的结构示意图;以及
[根据细则91更正 01.01.2019]
图6是本发明实施例三提供的计算设备的结构示意图。
图6是本发明实施例三提供的计算设备的结构示意图。
为了使本发明的目的、技术方案及优点更加清楚明白,以下结合附图及实施例,对本发明进行进一步详细说明。应当理解,此处所描述的具体实施例仅仅用以解释本发明,并不用于限定本发明。
以下结合具体实施例对本发明的具体实现进行详细描述:
实施例一:
图1示出了本发明实施例一提供的机器人柔顺性控制方法的实现流程,为了便于说明,仅示出了与本发明实施例相关的部分,详述如下:
在步骤S101中,获取示教运动的示教数据。
本发明实施例适用于对机器人的自动控制。机器人包括并不局限于机械臂、人形机器人等一系列带有关节、连杆等结构,并可实现伸缩、抓取等动作的机器人产品。其中,示教数据至少可包括示教运动的运动数据和交互力数据,因此,对示教运动的学习可包括动作学习和力学习(即,变刚度参数和变阻尼参数学习)。
在本发明实施例中,运动数据可包括机器人的预设点(例如,末端执行器)的位置数据和速度数据,或者运动数据可包括机器人的预设角(例如,关节角)的角度和角加速度,另外,运动数据还可包括其他可用于完整描述示教运动的一个或多个参数,本发明对此不进行限制。
作为示例,图2示出一种对机器人进行示教的示图,如图2所示,在示教时,示教者一只手抓握住机器人的末端执行器,在平面或空间中运动出一条轨迹,另一个手在末端施加示教力。机器人通过自带的运动捕捉系统和腕部装的六维力传感器采集示教数据。
例如,当运动数据包括机器人的预设采样点(例如,末端执行器或者末端等)的位置数据和速度数据时,按照时间间隔对末端的位置数据和交互力数据进行采样,从而获取一系列采样点数据
其中i=1,...,N
traj,N
traj表示示教运动轨迹的数量,k=1,...,N
i,N
i表示示教中采样点个数(每隔一个时间间隔采一次样),
即表示第i条轨迹的第k个采样点的末端位置,F
k第k个采样点的示教力大小。
作为另一示例,在示教时,示教者通过遥控器或示教器控制机器人进行示教操作,或者手把手示教。机器人根据示教操作记录示教数据。
作为又一示例,在示教时,示教者亲自完成示教运动任务。由机器人的运 动捕捉器或数据手套以及力传感器等设备根据示教运动采集示教数据。
优选地,如果运动数据包括机器人的预设点(例如,末端执行器)的位置数据和速度数据,则在获取示教运动的示教数据时,可首先获取与示教运动相关的位置数据、交互力数据和时间数据,然后根据位置数据和时间数据计算与示教运动相关的速度数据,从而得到示教运动的运动数据。
优选地,如果运动数据包括机器人的预设角(例如,关节角)的角度和角加速度,则在获取示教运动的示教数据时,可首先获取与示教运动相关的角度数据、交互力数据和时间数据,然后根据角度数据和时间数据计算与示教运动相关的角加速度数据,从而得到示教运动的运动数据。
在步骤S102中,根据示教数据中的运动数据计算示教运动的运动方程,并且同时根据示教数据中的交互力数据计算示教运动的变阻抗参数。
在本发明实施例中,变阻抗参数至少可包括变刚度参数和变阻尼参数。同时对运动轨迹和力进行学习,从而提高学习效果,进而提高控制结果的精确性。
优选地,在根据示教数据中的运动数据计算示教运动的运动方程时,可使用运动数据对预设的神经网络模型进行训练,得到示教运动的运动方程,并根据运动方程对神经网络模型进行在线更新,从而提高运动方程的计算效率,方便运动方程的后续使用,并且适应实时在线学习的需要,进而提高学习效果。
其中,优选地,在使用运动数据对预设的神经网络模型进行训练时,可以以逐一或逐块的方式对运动数据进行增量学习,得到示教运动的运动方程,从而提高运动方程的准确性,进而提高学习效果。
神经网络模型可以是支持向量机(Support Vector Machine,SVM)、在线序列超限学习机等能够增量在线学习的模型,也可以是其他的增量在线学习的模型,如增量支持向量机(ISVM)等,本发明对此不进行限定。其中,由于与其他在线学习算法相比,在线序列超限学习机具有学习速度快、泛化能力强、实现简单等特点,因此,优选地,神经网络模型是在线序列超限学习机,即使用运动数据对在线序列超限学习机进行训练,从而提高训练效率。
其中,在对在线序列超限学习机进行训练时,在运动方面,输入输出分别为采样点(例如,机器人末端执行器)的位置和速度(或者关节角的角度和角加速度),因此,在线序列超限学习机的输入和输出应该具有相同的维度,即具有相同的神经元个数d。如果考虑二维平面内的运动,d=2,如果考虑三维空间内的运动,d=3。
作为示例,图3示出在线序列超限学习机的示例性结构,如图3所示,假设在线序列超限学习机的隐藏层的激活函数为g,那么我们要学习的在线序列超限学习机(即,要学习的模型)可以表达为
其中,隐藏层神经元个数为
为隐藏层的偏置,
为隐藏层的权值,维度为
为输出层的权值,维度为
其中,虽然激活函数g一般选择S形函数(sigmoid函数)或双曲正切函数(tanh函数),也可使用的修改后的S形函数,例如,
但是,只要是满足
并且单调递增的连续、连续可微的函数都符合激活函数的要求,在此不做限制。
在线序列超限学习机的训练目标是要找到一组最佳的输出层权值
使用最小二乘法可得到
其中,H
+是矩阵H的Moore-Penrose广义逆矩阵。使用这种方法可不经过迭代求得输出层权值,在增加约束条件时,求解输出层权值的问题就变成了一个带约束的优化问题。
其中,在线序列超限学习机的训练过程包括一个初始的ELM批量学习过程和一个连续的贯序学习过程,具体如下:
作为示例,在根据示教数据中的交互力数据计算示教运动的变阻抗参数时,可根据交互力数据计算变刚度参数和变阻尼参数。
具体地,在计算变刚度参数时,令
表示采集到的交互力(F)与相应的时间(q)信息,其中是所获得扰动数据样本的个数。在q时刻的变刚度参数是由时间窗[q-(w-1),q]内的力信息计算得到。滑动时间窗的长度为w,窗内数据点的上下界分别用L
q,U
q表示,
在q时刻窗内数据点的个数为W
q=U
q-L
q+1,窗内力数据对应的的协方差矩阵为
其中,
由于协方差矩阵Σ
q是对称且正定的,故其能分解成如下形式Σ
q=PΛP-
1,其中,Λ是包含特征值
的对角阵。刚度矩阵K
q为
其中,
与特征值
成正比,表达式为
随着示教的进行,交互力的数据会不断的被收集,并根据时间信息将新的数据进行排序并取窗内的值进行刚度求解。例如当q+1时刻的数据进入时,协方差的在线更新为
其中,
在本发明实施例中,为保证学习出的模型具有稳定性,优选地,在根据交互力数据计算示教运动的变阻抗参数时,可根据预设的稳定性约束条件和交互力数据对预设的变阻抗模型进行训练,得到示教运动的变阻抗参数,并根据变阻抗参数对变阻抗模型进行更新,从而保证变阻抗控制的稳定性,避免出现机器人交互力过大而出现伤人的情况。
在步骤S103中,根据运动方程和变阻抗参数控制操作。
在本发明实施例中,在得到运动方程和变阻抗参数之后,可根据运动方程和变阻抗参数控制操作,从而控制机器人复现示教运动的运动轨迹和交互力。
在本发明实施例中,在预设的神经网络模型(例如,在线序列超限学习机)和变阻抗模型训练完成之后,可使用训练后的神经网络模型(例如,在线序列超限学习机)和变阻抗模型,从而控制机器人复现示教运动的运动轨迹和交互力。
作为示例,图4示出机器人柔顺性控制的示教学习和复现的示例性示图,如图4所示,示教者一只手抓握住机器人进行示教,机器人采集示教运动的轨 迹信息
和力信息F
q,然后根据轨迹信息
进行动作学习得到f(·),并且根据力信息F
q进行变刚度参数和变阻尼参数学习得到{B
q,K
q},最后根据f(·)产生运动,并根据{B
q,K
q}进行变阻抗控制,从而控制机器人复现示教运动的运动轨迹和交互力。
在本发明实施例中,通过获取示教运动的示教数据,根据示教数据中的运动数据计算示教运动的运动方程,并且同时根据示教数据中的交互力数据计算示教运动的变阻抗参数,根据运动方程和变阻抗参数控制操作,从而减少了机器人柔顺性控制过程中的手动编程,降低了机器人的使用门槛,提高了机器人控制的柔顺性和精确性,进而提高了机器人的泛化能力、智能化程度和控制效果。
实施例二:
图5示出了本发明实施例二提供的机器人柔顺性控制装置的结构,为了便于说明,仅示出了与本发明实施例相关的部分,其中包括:数据获取单元51、参数计算单元52和操作控制单元53。
数据获取单元51,用于获取示教运动的示教数据,其中,示教数据至少包括示教运动的运动数据和交互力数据。
在本发明实施例中,对示教运动的学习可包括动作学习和力学习(即,变刚度参数和变阻尼参数学习)。
在本发明实施例中,运动数据可包括机器人的预设点(例如,末端执行器)的位置数据和速度数据,或者运动数据可包括机器人的预设角(例如,关节角)的角度和角加速度,另外,运动数据还可包括其他可用于完整描述示教运动的一个或多个参数,本发明对此不进行限制。
因此,优选地,数据获取单元51可包括:
第一获取单元,用于获取与示教运动相关的位置数据、交互力数据和时间数据;以及
第一计算单元,用于根据位置数据和时间数据计算运动数据。
具体地,根据位置数据和时间数据计算与示教运动相关的速度数据,从而得到示教运动的运动数据。
优选地,数据获取单元51还可包括:
第二获取单元,用于获取与示教运动相关的角度数据、交互力数据和时间数据;以及
第二计算单元,用于根据角度数据和时间数据计算运动数据。
具体地,可根据角度数据和时间数据计算与示教运动相关的角加速度数据,从而得到示教运动的运动数据。
参数计算单元52,用于根据示教数据中的运动数据计算示教运动的运动方程,并且同时根据示教数据中的交互力数据计算示教运动的变阻抗参数,其中,变阻抗参数至少包括变刚度参数和变阻尼参数。
在本发明实施例中,同时对运动轨迹和力进行学习,从而提高学习效果,进而提高控制结果的精确性。
优选地,参数计算单元52可包括:
第一训练单元,用于使用运动数据对预设的神经网络模型进行训练,得到示教运动的运动方程,并根据运动方程对神经网络模型进行在线更新,从而提高运动方程的计算效率,方便运动方程的后续使用,并且适应实时在线学习的需要,进而提高学习效果。
其中,优选地,模型训练单元可包括:
增量学习单元,用于以逐一或逐块的方式对运动数据进行增量学习,得到示教运动的运动方程,从而提高运动方程的准确性,进而提高学习效果。
其中,优选地,神经网络模型是在线序列超限学习机。
优选地,参数计算单元52可包括:
第二训练单元,用于根据预设的稳定性约束条件和交互力数据对预设的变阻抗模型进行训练,得到示教运动的变阻抗参数,并根据变阻抗参数对变阻抗模型进行更新,从而保证变阻抗控制的稳定性,避免出现机器人交互力过大而 出现伤人的情况。
操作控制单元53,用于根据运动方程和变阻抗参数控制操作。
在本发明实施例中,通过数据获取单元51获取示教运动的示教数据,通过参数计算单元52根据示教数据中的运动数据计算示教运动的运动方程,并且同时根据示教数据中的交互力数据计算示教运动的变阻抗参数,通过操作控制单元53根据运动方程和变阻抗参数控制操作,从而减少了机器人柔顺性控制过程中的手动编程,降低了机器人的使用门槛,提高了机器人控制的柔顺性和精确性,进而提高了机器人的泛化能力、智能化程度和控制效果。
在本发明实施例中,机器人柔顺性控制装置的各单元可由相应的硬件或软件单元实现,各单元可以为独立的软、硬件单元,也可以集成为一个软、硬件单元,在此不用以限制本发明。
实施例三:
图6示出了本发明实施例四提供的计算设备的结构,为了便于说明,仅示出了与本发明实施例相关的部分。
本发明实施例的计算设备6包括处理器60、存储器61以及存储在存储器61中并可在处理器60上运行的计算机程序62。该处理器60执行计算机程序62时实现上述各个机器人柔顺性控制方法实施例中的步骤,例如图1所示的步骤S101至S103。或者,处理器60执行计算机程序62时实现上述各装置实施例中各单元的功能,例如,图5所示单元51至53的功能。
在本发明实施例中,该处理器60执行计算机程序62时实现上述各个机器人柔顺性控制方法实施例中的步骤时,获取示教运动的示教数据,根据示教数据中的运动数据计算示教运动的运动方程,并且同时根据示教数据中的交互力数据计算示教运动的变阻抗参数,根据运动方程和变阻抗参数控制操作,从而减少了机器人柔顺性控制过程中的手动编程,降低了机器人的使用门槛,提高了机器人控制的柔顺性和精确性,进而提高了机器人的泛化能力、智能化程度和控制效果。
该计算设备6中处理器60在执行计算机程序62时实现的步骤具体可参考实施例一中方法的描述,在此不再赘述。
实施例五:
在本发明实施例中,提供了一种计算机可读存储介质,该计算机可读存储介质存储有计算机程序,该计算机程序被处理器执行时实现上述各个机器人柔顺性控制方法实施例中的步骤,例如,图1所示的步骤S101至S103。或者,该计算机程序被处理器执行时实现上述各装置实施例中各单元的功能,例如,图5所示单元51至53的功能。
在本发明实施例中,获取示教运动的示教数据,根据示教数据中的运动数据计算示教运动的运动方程,并且同时根据示教数据中的交互力数据计算示教运动的变阻抗参数,根据运动方程和变阻抗参数控制操作,从而减少了机器人柔顺性控制过程中的手动编程,降低了机器人的使用门槛,提高了机器人控制的柔顺性和精确性,进而提高了机器人的泛化能力、智能化程度和控制效果。该计算机程序被处理器执行时实现的机器人柔顺性控制方法进一步可参考前述方法实施例中步骤的描述,在此不再赘述。
本发明实施例的计算机可读存储介质可以包括能够携带计算机程序代码的任何实体或装置、记录介质,例如,ROM/RAM、磁盘、光盘、闪存等存储器。
以上所述仅为本发明的较佳实施例而已,并不用以限制本发明,凡在本发明的精神和原则之内所作的任何修改、等同替换和改进等,均应包含在本发明的保护范围之内。
Claims (12)
- 一种机器人柔顺性控制方法,其特征在于,所述方法包括下述步骤:获取示教运动的示教数据,其中,所述示教数据至少包括所述示教运动的运动数据和交互力数据;根据所述示教数据中的运动数据计算所述示教运动的运动方程,并且同时根据所述示教数据中的交互力数据计算所述示教运动的变阻抗参数,其中,所述变阻抗参数至少包括变刚度参数和变阻尼参数;根据所述运动方程和所述变阻抗参数控制操作。
- 如权利要求1所述的方法,其特征在于,获取示教运动的示教数据的步骤,包括:获取与示教运动相关的位置数据、交互力数据和时间数据;根据所述位置数据和所述时间数据计算所述运动数据,其中,所述运动数据中包括速度数据。
- 如权利要求1所述的方法,其特征在于,根据所述示教数据中的运动数据计算所述示教运动的运动方程的步骤,包括:使用所述运动数据对预设的神经网络模型进行训练,得到所述示教运动的运动方程,并根据所述运动方程对所述神经网络模型进行在线更新,其中,使用所述运动数据对预设的神经网络模型进行训练的步骤,包括:以逐一或逐块的方式对所述运动数据进行增量学习,得到所述示教运动的运动方程。
- 如权利要求4所述的方法,其特征在于,所述神经网络模型包括在线序列超限学习机。
- 如权利要求1所述的方法,其特征在于,根据所述示教数据中的交互力数据计算所述示教运动的变阻抗参数的步骤,包括:根据预设的稳定性约束条件和所述交互力数据对预设的变阻抗模型进行训练,得到所述示教运动的变阻抗参数,并根据所述变阻抗参数对所述变阻抗模 型进行更新。
- 一种机器人柔顺性控制装置,其特征在于,所述装置包括:数据获取单元,用于获取示教运动的示教数据,其中,所述示教数据至少包括所述示教运动的运动数据和交互力数据;参数计算单元,用于根据所述示教数据中的运动数据计算所述示教运动的运动方程,并且同时根据所述示教数据中的交互力数据计算所述示教运动的变阻抗参数,其中,所述变阻抗参数至少包括变刚度参数和变阻尼参数;以及操作控制单元,用于根据所述运动方程和所述变阻抗参数控制操作。
- 如权利要求6所述的装置,其特征在于,所述数据获取单元包括:第一获取单元,用于获取与示教运动相关的位置数据、交互力数据和时间数据;以及第一计算单元,用于根据所述位置数据和所述时间数据计算所述运动数据,其中,所述运动数据中包括速度数据。
- 如权利要求6所述的装置,其特征在于,所述参数计算单元包括:第一训练单元,用于使用所述运动数据对预设的神经网络模型进行训练,得到所述示教运动的运动方程,并根据所述运动方程对所述神经网络模型进行在线更新,其中,所述模型训练单元包括:增量学习单元,用于以逐一或逐块的方式对所述运动数据进行增量学习,得到所述示教运动的运动方程。
- 如权利要求8所述的装置,其特征在于,所述神经网络模型包括在线序列超限学习机。
- 如权利要求6所述的装置,其特征在于,所述参数计算单元包括:第二训练单元,用于根据预设的稳定性约束条件和所述交互力数据对预设的变阻抗模型进行训练,得到所述示教运动的变阻抗参数,并根据所述变阻抗参数对所述变阻抗模型进行更新。
- 一种计算设备,包括存储器、处理器以及存储在所述存储器中并可在所述处理器上运行的计算机程序,其特征在于,所述处理器执行所述计算机程序时实现如权利要求1至5任一项所述方法的步骤。
- 一种计算机可读存储介质,所述计算机可读存储介质存储有计算机程序,其特征在于,所述计算机程序被处理器执行时实现如权利要求1至5任一项所述方法的步骤。
Priority Applications (1)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| PCT/CN2018/121338 WO2020118730A1 (zh) | 2018-12-14 | 2018-12-14 | 机器人柔顺性控制方法、装置、设备及存储介质 |
Applications Claiming Priority (1)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| PCT/CN2018/121338 WO2020118730A1 (zh) | 2018-12-14 | 2018-12-14 | 机器人柔顺性控制方法、装置、设备及存储介质 |
Publications (1)
| Publication Number | Publication Date |
|---|---|
| WO2020118730A1 true WO2020118730A1 (zh) | 2020-06-18 |
Family
ID=71075313
Family Applications (1)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| PCT/CN2018/121338 Ceased WO2020118730A1 (zh) | 2018-12-14 | 2018-12-14 | 机器人柔顺性控制方法、装置、设备及存储介质 |
Country Status (1)
| Country | Link |
|---|---|
| WO (1) | WO2020118730A1 (zh) |
Cited By (8)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| CN114366552A (zh) * | 2021-12-23 | 2022-04-19 | 威海经济技术开发区天智创新技术研究院 | 一种上肢康复训练外骨骼控制方法及系统 |
| CN114571444A (zh) * | 2022-03-01 | 2022-06-03 | 中南大学 | 一种基于Q-learning的扒渣机器人阻抗控制方法 |
| CN114800489A (zh) * | 2022-03-22 | 2022-07-29 | 华南理工大学 | 基于确定学习与复合学习联合的机械臂柔顺控制方法、存储介质及机器人 |
| CN114848391A (zh) * | 2022-04-28 | 2022-08-05 | 北京邮电大学 | 一种下肢康复机器人柔顺控制方法 |
| CN115421387A (zh) * | 2022-09-22 | 2022-12-02 | 中国科学院自动化研究所 | 一种基于逆强化学习的可变阻抗控制系统及控制方法 |
| WO2023124346A1 (zh) * | 2021-12-28 | 2023-07-06 | 广东省科学院智能制造研究所 | 一种协作机器人可变刚度运动技能学习与调控方法及系统 |
| CN116652938A (zh) * | 2023-05-16 | 2023-08-29 | 中国科学院长春光学精密机械与物理研究所 | 类人化手臂刚度与运动技能的机械臂学习方法、设备 |
| CN119369386A (zh) * | 2024-10-21 | 2025-01-28 | 深圳市大族机器人有限公司 | 机器人自适应阻抗控制方法、装置、计算机设备、可读存储介质和程序产品 |
Citations (5)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| EP0416123A1 (en) * | 1989-03-20 | 1991-03-13 | Fanuc Ltd. | Manual intervention method for industrial robot |
| JP2002120174A (ja) * | 2000-10-11 | 2002-04-23 | Sony Corp | オーサリング・システム及びオーサリング方法、並びに記憶媒体 |
| CN101909829A (zh) * | 2008-02-28 | 2010-12-08 | 松下电器产业株式会社 | 机器人手臂的控制装置及控制方法、机器人、机器人手臂的控制程序、及机器人手臂控制用集成电子电路 |
| CN104259991A (zh) * | 2014-09-01 | 2015-01-07 | 沈阳远大科技园有限公司 | 基于可变刚度柔性机构的力控制模块 |
| CN206484559U (zh) * | 2016-11-08 | 2017-09-12 | 武汉海默机器人有限公司 | 基于三维力传感器的机器人示教系统 |
-
2018
- 2018-12-14 WO PCT/CN2018/121338 patent/WO2020118730A1/zh not_active Ceased
Patent Citations (5)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| EP0416123A1 (en) * | 1989-03-20 | 1991-03-13 | Fanuc Ltd. | Manual intervention method for industrial robot |
| JP2002120174A (ja) * | 2000-10-11 | 2002-04-23 | Sony Corp | オーサリング・システム及びオーサリング方法、並びに記憶媒体 |
| CN101909829A (zh) * | 2008-02-28 | 2010-12-08 | 松下电器产业株式会社 | 机器人手臂的控制装置及控制方法、机器人、机器人手臂的控制程序、及机器人手臂控制用集成电子电路 |
| CN104259991A (zh) * | 2014-09-01 | 2015-01-07 | 沈阳远大科技园有限公司 | 基于可变刚度柔性机构的力控制模块 |
| CN206484559U (zh) * | 2016-11-08 | 2017-09-12 | 武汉海默机器人有限公司 | 基于三维力传感器的机器人示教系统 |
Cited By (11)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| CN114366552A (zh) * | 2021-12-23 | 2022-04-19 | 威海经济技术开发区天智创新技术研究院 | 一种上肢康复训练外骨骼控制方法及系统 |
| CN114366552B (zh) * | 2021-12-23 | 2023-10-24 | 威海经济技术开发区天智创新技术研究院 | 一种上肢康复训练外骨骼控制方法及系统 |
| WO2023124346A1 (zh) * | 2021-12-28 | 2023-07-06 | 广东省科学院智能制造研究所 | 一种协作机器人可变刚度运动技能学习与调控方法及系统 |
| CN114571444A (zh) * | 2022-03-01 | 2022-06-03 | 中南大学 | 一种基于Q-learning的扒渣机器人阻抗控制方法 |
| CN114800489A (zh) * | 2022-03-22 | 2022-07-29 | 华南理工大学 | 基于确定学习与复合学习联合的机械臂柔顺控制方法、存储介质及机器人 |
| CN114800489B (zh) * | 2022-03-22 | 2023-06-20 | 华南理工大学 | 基于确定学习与复合学习联合的机械臂柔顺控制方法、存储介质及机器人 |
| CN114848391A (zh) * | 2022-04-28 | 2022-08-05 | 北京邮电大学 | 一种下肢康复机器人柔顺控制方法 |
| CN114848391B (zh) * | 2022-04-28 | 2023-08-29 | 北京邮电大学 | 一种下肢康复机器人柔顺控制方法 |
| CN115421387A (zh) * | 2022-09-22 | 2022-12-02 | 中国科学院自动化研究所 | 一种基于逆强化学习的可变阻抗控制系统及控制方法 |
| CN116652938A (zh) * | 2023-05-16 | 2023-08-29 | 中国科学院长春光学精密机械与物理研究所 | 类人化手臂刚度与运动技能的机械臂学习方法、设备 |
| CN119369386A (zh) * | 2024-10-21 | 2025-01-28 | 深圳市大族机器人有限公司 | 机器人自适应阻抗控制方法、装置、计算机设备、可读存储介质和程序产品 |
Similar Documents
| Publication | Publication Date | Title |
|---|---|---|
| CN109702740B (zh) | 机器人柔顺性控制方法、装置、设备及存储介质 | |
| Yu et al. | Global model learning for large deformation control of elastic deformable linear objects: An efficient and adaptive approach | |
| WO2020118730A1 (zh) | 机器人柔顺性控制方法、装置、设备及存储介质 | |
| CN108115681B (zh) | 机器人的模仿学习方法、装置、机器人及存储介质 | |
| Gu et al. | Deep reinforcement learning for robotic manipulation with asynchronous off-policy updates | |
| CN114310888B (zh) | 一种协作机器人可变刚度运动技能学习与调控方法及系统 | |
| Wang et al. | A framework of hybrid force/motion skills learning for robots | |
| Guenter et al. | Reinforcement learning for imitating constrained reaching movements | |
| Calinon et al. | Learning and reproduction of gestures by imitation | |
| US11565412B2 (en) | Generating a robot control policy from demonstrations collected via kinesthetic teaching of a robot | |
| Billard et al. | Discovering optimal imitation strategies | |
| Stulp et al. | Model-free reinforcement learning of impedance control in stochastic environments | |
| Neumann et al. | Neural learning of stable dynamical systems based on data-driven Lyapunov candidates | |
| Mitrovic et al. | Adaptive optimal feedback control with learned internal dynamics models | |
| Khadivar et al. | Adaptive fingers coordination for robust grasp and in-hand manipulation under disturbances and unknown dynamics | |
| CN111872934A (zh) | 一种基于隐半马尔可夫模型的机械臂控制方法及系统 | |
| Yin et al. | Learning nonlinear dynamical system for movement primitives | |
| Ewerton et al. | Incremental imitation learning of context-dependent motor skills | |
| Liu et al. | Demonstration learning and generalization of robotic motor skills based on wearable motion tracking sensors | |
| CN117325182B (zh) | 一种融合多自适应调控机制的机器人柔顺控制方法 | |
| CN112405542B (zh) | 基于脑启发多任务学习的肌肉骨骼机器人控制方法及系统 | |
| Totsila et al. | Sensorimotor learning with stability guarantees via autonomous neural dynamic policies | |
| Kawaharazuka et al. | Adaptive robotic tool-tip control learning considering online changes in grasping state | |
| WO2019095108A1 (zh) | 机器人的模仿学习方法、装置、机器人及存储介质 | |
| Dong et al. | Robot learning from multiple demonstrations based on generalized gaussian mixture model |
Legal Events
| Date | Code | Title | Description |
|---|---|---|---|
| 121 | Ep: the epo has been informed by wipo that ep was designated in this application |
Ref document number: 18943108 Country of ref document: EP Kind code of ref document: A1 |
|
| NENP | Non-entry into the national phase |
Ref country code: DE |
|
| 32PN | Ep: public notification in the ep bulletin as address of the adressee cannot be established |
Free format text: NOTING OF LOSS OF RIGHTS PURSUANT TO RULE 112(1) EPC (EPO FORM 1205 DATED 04.11.2021) |
|
| 122 | Ep: pct application non-entry in european phase |
Ref document number: 18943108 Country of ref document: EP Kind code of ref document: A1 |







