Search arXivSearch

arXiv · 2010.03814

A novel control mode of bionic morphing tail based on deep reinforcement learning

Abstract

In the field of fixed wing aircraft, many morphing technologies have been applied to the wing, such as adaptive airfoil, variable span aircraft, variable swept angle aircraft, etc., but few are aimed at the tail. The traditional fixed wing tail includes horizontal and vertical tail. Inspired by the bird tail, this paper will introduce a new bionic tail. The tail has a novel control mode, which has multiple control variables. Compared with the traditional fixed wing tail, it adds the area control and rotation control around the longitudinal symmetry axis, so it can control the pitch and yaw of the aircraft at the same time. When the area of the tail changes, the maneuverability and stability of the aircraft can be changed, and the aerodynamic efficiency of the aircraft can also be improved. The aircraft with morphing ability is often difficult to establish accurate mathematical model, because the model has a strong nonlinear, model-based control method is difficult to deal with the strong nonlinear aircraft. In recent years, with the rapid development of artificial intelligence technology, learning based control methods are also brilliant, in which the deep reinforcement learning algorithm can be a good solution to the control object which is difficult to establish model. In this paper, the model-free control algorithm PPO is used to control the tail, and the traditional PID is used to control the aileron and throttle. After training in simulation, the tail shows excellent attitude control ability.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Liming Zheng, Zhou Zhou, Pengbo Sun, Zhilin Zhang, Rui Wang. 2020-10-08. A novel control mode of bionic morphing tail based on deep reinforcement learning. https://arxiv.org/abs/2010.03814

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Mathematical modeling on peristaltic flow of a Prandtl fluid with effects of slip conditions and inclined magnetic field

The manuscript provides a description of a theoretical analysis of a non-Newtonian Prandtl fluid subject to peristaltic flow through an inclined asymmetric channel. We explore the effect of an inclined magnetic field on the peristaltic flow. This is relevant for applications involving fluid flow in narrow, inclined (tilted) tubes similar to blood vessels or the digestive system. The model also includes thermodynamic aspects such as heat diffusion (the Soret effect) and viscous dissipation resulting from wall-fluid slip conditions, which may help optimize medical devices such as lab-on-a-chip systems and dialysis machines. In this study, the concentration of a generic chemical, temperature, and fluid velocity are taken into account through mass, heat, and momentum balances, respectively. The solution is approximated using numerical techniques suitable for long wavelengths (low frequency) and low Reynolds numbers. The study also discusses trapping phenomena, which are crucial from a clinical point of view. The developed insights can improve the understanding of physiological flows in the gastrointestinal tract and blood vessels. By understanding how the fluid moves and how particles are trapped, these insights may contribute to the design of improved medical pumps and artificial organs. Graphical visualizations are provided for the fluid velocity profile, temperature distribution, and concentration of a generic chemical. Furthermore, the numerical results are validated through comparison with a closed-form solution from a benchmark problem.

physics.flu-dyn

Discovery of a dispersion model at high Peclet numbers

Peclet number characterises the transition from classical Taylor-Aris dispersion to convection-dominated longitudinal solute transport, with the classical model becoming inadequate at extremely high radial Peclet number $Pe_r$. We develop a novel explicit-closure one-dimensional (1-D) effective dispersion model for this high-$Pe_r$ regime by introducing two closure coefficients, $θ_u$ and $θ_d$, whose functional structures are identified using low-frequency transfer-function matching and a modified Kolmogorov-Arnold network (KAN). The resulting model captures the transition from classical Taylor-Aris dispersion at low $Pe_r$ to convection-dominated dispersion at high $Pe_r$. Analysis reveals that, in the high-$Pe_r$ regime, axial transport is redistributed between the effective convection flux and the dispersive flux, resulting in a reduced macroscopic convection velocity. Numerical validation demonstrates close agreement with the convection-diffusion model over the investigated high-$Pe_r$ conditions, while the classical Taylor-Aris model exhibits substantial deviations. Application of the proposed model to averaged flow velocity inversion further demonstrates improved velocity estimation, particularly in the high-$Pe_r$ regime. These results highlight the importance of accounting for non-classical dispersion for reliable contrast-agent-based arterial blood flow velocimetry and provide new insight into high-$Pe_r$ mass transport.

physics.flu-dyn

Optimization of fluid mixing by reinforcement learning using limit cycles of a dynamical system

We propose a method to overcome the difficulties encountered when applying reinforcement learning to fluid mixing processes. The proposed method has two main features: (i) it does not require detailed measurements of the flow state, and (ii) by effectively exploiting a stable limit cycle of a two-dimensional dynamical system (the Li'enard system), it can stably perform optimization without imposing explicit constraints on the control parameters. As an illustrative example, we optimize a process in which a fluid contained in a cylindrical vessel is mixed by periodically rotating the vessel. The resulting optimal vessel motion is physically reasonable: it reverses its direction of rotation before a solid-body rotation state is established. Furthermore, even when the fluid viscosity increases with time during the mixing process, the method can continuously adapt the control parameters to the changing viscosity.

physics.flu-dyn