-
Pb Substitution Effects on Lattice and Electronic System of the BiS2-based Superconductors La(O F)BiS2
Authors:
Miku Sasaki,
Kotaro Inada,
Fumito Mori,
Takaaki Hirase,
Haruki Yamada,
Shota Shimoyama,
Yasushi Nakamura,
Tetta Nakamura,
Yoshiyuki Shibayama,
Naoki Momono
Abstract:
We examined the effect of Pb substitution in the layered superconductor LaO0.5F0.5Bi1-xPbxS2 (x=0~0.15) through the measurements of the resistivity, thermal expansion, specific heat, and Seebeck coefficient. These transport and thermal properties show anomalies at certain temperatures (T*) for x${\geq}$0.08. The large thermal expansion anomalies, specific heat anomalies, and the existence of hyste…
▽ More
We examined the effect of Pb substitution in the layered superconductor LaO0.5F0.5Bi1-xPbxS2 (x=0~0.15) through the measurements of the resistivity, thermal expansion, specific heat, and Seebeck coefficient. These transport and thermal properties show anomalies at certain temperatures (T*) for x${\geq}$0.08. The large thermal expansion anomalies, specific heat anomalies, and the existence of hystereses in the above measurements indicate a first-order structural phase transition at T*. Additionally, the Seebeck coefficient indicates that the anomalies at T* are related not only to the lattice system, but also to the electronic system. Superconductivity is not observed above 2 K at x=0.08, which is around the phase boundary where T* vanishes. The suppression of superconductivity around the structural phase boundary suggests a close relationship between the lattice and superconductivity.
△ Less
Submitted 5 June, 2024; v1 submitted 3 June, 2024;
originally announced June 2024.
-
Transformation of $p$-gradient flows to $p'$-gradient flows in metric spaces
Authors:
Sho Shimoyama
Abstract:
We explicitly construct parameter transformations between gradient flows in metric spaces, called curves of maximal slope, having different exponents when the associated function satisfies a suitable convexity condition. These transformations induce the uniqueness of gradient flows for all exponents under a natural assumption which is satisfied in many examples. We also prove the regularizing effe…
▽ More
We explicitly construct parameter transformations between gradient flows in metric spaces, called curves of maximal slope, having different exponents when the associated function satisfies a suitable convexity condition. These transformations induce the uniqueness of gradient flows for all exponents under a natural assumption which is satisfied in many examples. We also prove the regularizing effects of gradient flows. To establish these results, we directly deal with gradient flows instead of using variational discrete approximations which are often used in the study of gradient flows.
△ Less
Submitted 3 April, 2024;
originally announced April 2024.
-
Why Guided Dialog Policy Learning performs well? Understanding the role of adversarial learning and its alternative
Authors:
Sho Shimoyama,
Tetsuro Morimura,
Kenshi Abe,
Toda Takamichi,
Yuta Tomomatsu,
Masakazu Sugiyama,
Asahi Hentona,
Yuuki Azuma,
Hirotaka Ninomiya
Abstract:
Dialog policies, which determine a system's action based on the current state at each dialog turn, are crucial to the success of the dialog. In recent years, reinforcement learning (RL) has emerged as a promising option for dialog policy learning (DPL). In RL-based DPL, dialog policies are updated according to rewards. The manual construction of fine-grained rewards, such as state-action-based one…
▽ More
Dialog policies, which determine a system's action based on the current state at each dialog turn, are crucial to the success of the dialog. In recent years, reinforcement learning (RL) has emerged as a promising option for dialog policy learning (DPL). In RL-based DPL, dialog policies are updated according to rewards. The manual construction of fine-grained rewards, such as state-action-based ones, to effectively guide the dialog policy is challenging in multi-domain task-oriented dialog scenarios with numerous state-action pair combinations. One way to estimate rewards from collected data is to train the reward estimator and dialog policy simultaneously using adversarial learning (AL). Although this method has demonstrated superior performance experimentally, it is fraught with the inherent problems of AL, such as mode collapse. This paper first identifies the role of AL in DPL through detailed analyses of the objective functions of dialog policy and reward estimator. Next, based on these analyses, we propose a method that eliminates AL from reward estimation and DPL while retaining its advantages. We evaluate our method using MultiWOZ, a multi-domain task-oriented dialog corpus.
△ Less
Submitted 13 July, 2023;
originally announced July 2023.